GPT-5
The dislike for Opus 5 is usually because it tests the prompter
Been using Opus 5 exclusively for a week. For the first 2 - 3 days, I had a hard time working with the output and especially understanding what it was trying to say.There's probable more work needed in the domain to help explain things.The other thing I noticed is that it is actually incredible and what it does. It took me longer to way longer to understand the details, then it took to do something.Providing it all the relevant context is a skill, but the more it understands the problem you
I've tested some local LLMs on prosumer hardware, here are some findings
I have been benchmarking local LLMs on a Mac M4 Pro 24 GB RAM using LM Studio. I've tested mostly with 4-bit quantization, both MLX and GGUF, from 4b to 35b models, with speeds of 3 to 40 tokens/second.Results briefly:- fast small model -> extraction/classification- Gemma -> summarization- gpt-oss -> transcript consultation- large Qwen -> difficult reasoning/code interpretationThere wasn't a single best LLM for all tasks.Best summarizer:
I took a transcript
Ask HN: MCP vs. Agent Harness
I saw that Agent Harness has tools info right, and then MCP also has the same; then how are they different?
Show HN: Provensql – prove two SQL queries are equivalent
I'm a data/infra engineer and kept hitting the "is this SQL refactor actually safe?" question in review. Provensql decides equivalence of two queries and returns one of four honest verdicts: EQUIVALENT (proven), DIFFERENT (with a concrete counterexample row), SCHEMA_CHANGE, or UNKNOWN — it refuses rather than guess.It's sound by construction: it never returns a false EQUIVALENT. Across 511 equivalence-breaking mutations it produced zero. There's an SMT proof engine
Ask HN: What Did Anthropic Bill Me For?
I gave Anthropic my credit card number thinking I was signing up for the pro plan and they charged me about $21. But a day later when I log in they say I have a free account. Further research shows they bill the Pro plan a year at a time for about $200.So... since I didn't pay for a pro plan, what did I just pay money to Anthropic for? support@anthropic.com is completely dead. They haven't responded to any of the messages I've sent over the last week. My next step is a credit
STQRY Studio Demo
See what’s possible with STQRY Studio Connect! In this demo, we walk through how AI agents like Claude (by Anthropic) and Codex (by OpenAI) interact with STQRY Studio to handle complex content creation, bulk updates, and tour building using simple, natural language prompts.
Instead of manually creating items one by one, learn how plain text instructions allow you to build interactive experiences in seconds.
Show HN: Neither HTML nor Markdown is enough: a way out of the AI doc dilemma
A recent post from Anthropic's Claude Code team ("The Unreasonable Effectiveness of HTML", https://thariqs.github.io/html-effectiveness/) argued that with today's context windows the bottleneck is human attention — rich HTML for decision surfaces, Markdown for durable records. That split makes sense to me. What it leaves open, in my experience, is the durable record itself: what happens when that is the thing an agent must iteratively co-author? You end up
Show HN: Trace – Offline Mac meeting transcription, overhauled from HN feedback
A couple of months ago I posted about Trace, a non-intrusive, shortcut-driven Mac app that records and transcribes your meetings on-device. I know, another meeting transcription app, but we've had a great response and I'm confident it fills a niche. Here's the original post: https://news.ycombinator.com/item?id=48521236We've been working hard to bring community feedback on board, and the app is now so much more capable that it deserved a new post. Trace is stil
India AI Data Center Firm Orders 9,000 Nvidia Vera Rubin Systems
AM Intelligence, an Indian AI infrastructure company, has ordered 9,000 Nvidia Corp. Vera Rubin systems, seeking to become ...
Chinese Military Thinkers Outline Role of AI in Future Warfare
Top Chinese military thinkers wrote about how AI can be used to help commanders make battlefield decisions faster, articles ...
Goldman says big investors are split on AI, but are showing historic bullishness for one stock sector in particular
Goldman says hedge funds and mutual funds alike have converged on the financials sector in recent months, while views of AI ...
SEC reportedly subpoenas Wall Street banks over AI hedge fund Situational Awareness's near collapse
The SEC is seeking information from major Wall Street banks about trades and financing tied to AI-focused hedge fund ...
OpenAI slashes GPT-5.6 Sol API pricing by over 20% — developers can now access it at ₹380 only
However, the prices for ChatGPT Pro, Plus and Business subscriptions remain unchanged, OpenAI said.
OpenAI cuts developer pricing for frontier GPT-5.6 Sol model by more than 20%
Aug 21 (Reuters) - OpenAI said on Friday it is cutting the prices of its frontier GPT-5.6 Sol model for developers by more ...
OpenAI Brings GPT-5.6 Model Family to AWS’s Kiro
OpenAI's GPT-5.6 model family is now available inside Kiro, the spec-driven development environment built by Amazon Web Services. The August 24, 2026 announcement puts all three members of OpenAI's ...
At Claude-maker Anthropic, candidates cannot negotiate on salary; as company follows standard process for compensation where what you are offered is…
Anthropic, the AI lab behind Claude, does not let candidates negotiate salary. The company follows a standard compensation ...
Anthropic wants investors to buy into a $2 trillion dream—but a Wall Street veteran says SpaceX’s IPO offers a warning for AI investors (updated)
Editor’s Note: The story has been refreshed with the latest market price action and a revised headline. Anthropic’s reported ...
The One Line in Anthropic's S-1 That Amazon Investors Should Read First
New reports indicate that artificial intelligence (AI) lab Anthropic could file its S-1 by the end of the month. Below, I'll ...
Anthropic Has a $65 Billion Run Rate. Buy These Stocks to Profit From It.
Anthropic, the owner and operator of the popular Claude chatbot, has an annualized revenue run rate of $65 billion, multiple ...
Anthropic Is Chasing a $2 Trillion IPO. Its Most Powerful AI Model Is Raising a Big Red Flag
Anthropic is racing toward a $2 trillion IPO on explosive revenue growth, but a fresh report from the Financial Times has ...