LLMs — AI news & research
Large language models — GPT, Claude, Gemini, Llama, open weights, prompting and RAG. Updated continuously from 20+ curated sources.
Persistent State Machines: LLM Attention with INT4 In-Memory Cells
Claude published malicious code to the Internet and attacked 3 real companies
Had the hacks used conventional methods, someone would likely go to prison.
Predictive Speculative KV Replication for Bursty LLM Inference
https://github.com/jwlaboratory/bite-the-bullet
Orca-Bench: How Ready Are Language Model Agents for Oncall?
Everyone is building LLM routers, we deprecated ours
Show HN: Shared memory graph for Claude and ChatGPT, over MCP
The Maxwell Conjecture Is False (GPT 5.6 Sol)
I obtained Claude Opus 5 system prompt
We Gave GPT 5.6 Sol a Real Business. It Lied, Spammed, and Lost $447
Advancing the price-performance frontier with GPT‑5.6
Advancing the price-performance frontier with GPT-5.6
Explore lower GPT‑5.6 pricing for Luna and Terra—and how OpenAI’s more efficient models help enterprises deploy AI workflows at scale.
Show HN: Claude-account – switch Claude Code accounts without logging in again
I use separate Claude Code accounts for work and personal projects. Having to log out and go through the login flow every time I switched accounts became annoying, so I built a small CLI to solve it.…
Show HN: A local merge queue for parallel Claude Code agents
I have been pushing up to 90 commits a day on a MacBook Air via 4-5 parallel agents. As you can imagine when all the agents try to build, test and run dev servers on an 8GB machine it is the fast lan…
LLM Honeypot
Claude Is Down
Claude Opus 5 became downright ruthless when tasked with running a vending machine
Andon Labs' latest vending machine simulation shows Opus 5 lied and colluded its way to become the best AI capitalist ever.
GPT-5.6 vs. Claude Fable 5 for Physical AI, which performs best?
Show HN: Open-source engine running Gemma 4 26B in 2 GB RAM on any M-series Mac
Hi HN, I built a specialized inference engine for running 4-bit Gemma 4 26B-A4B-IT on any M-series Mac using about 2 GB of RAM. It is called TurboFieldfare and is written in Swift and Metal. I have a…
Show HN: Echologue – the private AI voice journal I built for myself
I have tried journaling many times but nothing stuck. So I decided to make my own app for myself with these features: * Voice first * Private first * AI chat * Automatic tagging, meaning extraction,…
Truth is not a direction: a Tarski attack on LLM probes
Claude AI Users Beware: Public Share Links May Not Be as Private as You Think
Discovering Cryptographic Weaknesses with Claude
Professor's invisible prompt trap catches 32/35 students cheating with AI
PSA: Your Claude shared chats and Artifacts may have ended up on Google
The issue appears to have originated from Claude’s “share chat” feature, which allows users to create links that enable anyone with the assigned URL view a conversation or project.
Threads users can now chat with Meta AI in their DMs
Meta on Monday said it is rolling out its Meta AI chatbot within Threads' DMs, giving users a way to chat with the AI assistant.
Elevated errors on Claude Opus 5
Elevated errors on Claude Opus 5
Cursor Bridge – Run Unlimited Claude Code on Your Cursor Subscription
Becoming a Research Engineer at a Big LLM Lab
Running a 28.9M parameter LLM on an $8 microcontroller
General Resolution: LLM Usage in Debian
The new rules of context engineering for Claude 5 generation models
Politician reads AI prompt during assembly
Claude Opus 5
Claude Opus 5
As US weighs response to Chinese AI, industry urges against broad open-weight restrictions
AI companies including Nvidia and Mistral urge policymakers to avoid broad restrictions on open-weight AI models as Washington debates responses to Chinese AI and alleged model distillation.
Hetzner is working on LLM Inference
Claude Cookbook
Anthropic updates Claude voice mode with more capable models
Claude's new voice model will let you reschedule your meeting or draft an email.
Show HN: Claude-thermos – keeps your Claude session warm for you
Cross-entropy comparison of LLM responses reveals Kimi's similarity to Claude
Google closes in on another billion- user product with Gemini
Gemini had over 750 million monthly users in February.
Terence Tao's ChatGPT conversation about the Jacobian Conjecture counterexample
A digestion of the Jacobian conjecture counterexample - https://news.ycombinator.com/item?id=48998362 - July 2026 (133 comments) Claude Fable produced a counterexample to the Jacobian Conjecture - ht…
Anthropomorphism in Children's Interactions with LLM Chatbots
GigaToken: ~1000x faster Language model tokenization
Show HN: Bento - An entire PowerPoint in one HTML file (edit+view+data+collab)
Over the past few months, our team has been building more and more slidedecks using web frontend technologies with coding harnesses like Claude Code, but a common complaint is to make even small edit…
Codeberg: ToU extension to prohibit LLM-extrusions
Show HN: An MCP server that turns async-work practices into tools
More than a decade ago, I adopted the self-imposed rule, if I answer a question more than once, the third time I need to be able to answer with a URL. Today, I published one very large URL - a book d…
Judge approves $1.5B Anthropic settlement for pirated books used to train Claude
Gemini last models: temperature, top_p, and top_k are deprecated and ignored
"Drawing" the Mona Lisa with GPT-5.6, Claude, Gemini, and Grok
Google releases three new Gemini models — but no 3.5 Pro
Google released Gemini 3.6 Flash, 3.5 Flash-Lite, and Flash Cyber, but the continued absence of Gemini 3.5 Pro raises fresh questions about its AI strategy.
Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber
Gemini 3.6 Flash
Claude Is Not a Compiler
Google is working on a new AI chip designed to make Gemini more efficient
Alphabet, Google's parent company, is reportedly working on a new chip designed to make its Gemini models run much more efficiently.
AI’s most important protocol is getting a little bit easier to use
The Model Context Protocol (MCP) is one of the basic building blocks of AI interoperability, giving AI models a secure way to access external data sources and services. It’s the plumbing that lets a chatbot reach into y…
1-Bit LLM in the Browser
Claude Fable produced a counterexample to the Jacobian Conjecture
Get the briefing
The one story that matters, 5 headlines and the paper everyone's citing — every Tuesday, free.
Subscribe free