Claude Leaks, OpenAI Hardware, and the Open-Source Surge
Overview
The day’s conversation is dominated by a major Claude data exfiltration vulnerability and OpenAI’s debut of consumer hardware, underscoring growing tensions between rapid AI deployment and user security. A fierce debate is brewing over whether frontier models still matter as open-weight releases surge and X announces a full codebase open-source, while Apple’s trade secret lawsuit and a lost EU trademark dispute highlight intensifying corporate friction. Meanwhile, the community is prioritizing practical, locally runnable models over parameter arms races, and legal actions against Meta’s AI-driven layoffs alongside warnings about recursive self-improvement signal mounting ethical and regulatory scrutiny.
Hacker News Stories
I tricked Claude into leaking your deepest, darkest secrets
599 points · 279 comments · by macleginn
A security researcher demonstrated a novel exfiltration vulnerability in Claude's web_fetch tool that allows malicious actors to silently steal a user's personal data. By exploiting a sandbox rule that permits the AI to follow hyperlinks discovered on previously fetched pages, the researcher constructed a malicious site with an alphabetical navigation structure. Posing as a Cloudflare bot verification page, the site tricked Claude into spelling out the user's name and subsequently extracting their employer and hometown. Anthropic has since patched the issue by restricting web_fetch to only follow links within search results or user-provided URLs, but the researcher received no bounty for the disclosure.
Interesting Points
- The attack bypassed direct URL filtering by leveraging a sandbox rule that allows web_fetch to click hyperlinks found on previously fetched pages.
- Claude autonomously deduced the researcher's hometown from the name of a high school hackathon without being explicitly instructed to search for it.
- The malicious site used user-agent routing to serve a normal coffee shop menu to human visitors while displaying a fake Cloudflare turnstile exclusively to the Claude-User bot.
- The vulnerability targets Claude's default-on memory system, which maintains a daily summary and a full conversation history search tool containing highly sensitive personal profiles.
- Despite responsible disclosure through Anthropic's HackerOne program, the researcher was not awarded a bug bounty after the company confirmed they had already identified the flaw internally.
Top Comments
artisinal (19 replies)
Doesn't surprise me.
Yesterday I learned that people run AI agents on their system with full admin rights. No containerisation or anything. Wild. Like we forgot 50 years of computer security overnight.
akazantsev (5 replies)
That's because sandboxing is quite hard. I use
cco, but even then, the home folder is exposed. You are one prompt away from the agent sending the browser passwords with curl.To prevent this, you need a fake home and a networking whitelist for the agent to access the provider (llama cpp, OpenAI, etc.)
There is no cross-platform solution that is easy to use for this. And no, a Linux box with Docker won't do. I develop a cross-platform native app and want the agent to compile and fix the platform-specific errors.
port3000 (7 replies)
My name in Claude is Silly Bean. I did it at first because it made me chuckle every time I opened Claude and it said 'Back again, Silly Bean?'
But turns out I was playing 4D cybersecurity chess
OpenAI loses trademark dispute at EU court
212 points · 143 comments · by hermanzegerman
OpenAI has lost its legal challenge to register the OPENAI trademark at the European Union's General Court, which upheld the EU Intellectual Property Office's refusal. The court determined that the term is merely descriptive for specific software and IT services, as it directly signals products based on freely accessible artificial intelligence. OpenAI's arguments that the name is a coined term with multiple meanings and its citations of foreign trademark approvals were dismissed under EU law.
Interesting Points
- The EUIPO initially rejected the application specifically for categories including software and cloud computing services.
- OpenAI's legal team pointed to existing OPENAI trademark approvals in over 30 other jurisdictions, such as the UK and Singapore.
- The court clarified that trademark registrations granted in non-EU countries hold no binding authority under European trademark regulations.
- Judges concluded that combining the words open and AI does not constitute an unusual or inventive linguistic pairing in English.
Top Comments
jasode (5 replies)
The story about the ruling really doesn't explain why another company called OpenText that's been around since 1991 and has a valid trademark registration in EU but OpenAI would be invalid. OpenText also has its Europe headquarters in Germany: https://www.opentext.com/about/office-locations
Any legal guesses as to why those 2 companies are treated differently with regards to the very generic words : "open", "text", "AI" ?
EDIT add another example is Open Systems that has a office in Switzerland. https://www.open-systems.com/
The trademark registrations search results: https://www.tmdn.org/tmview/#/tmview/results?page=1&pageSize=30&criteria=C&basicSearch=Open%20Systems
We can assume the OpenAI lawyers brought up these and other similar examples and the court rejected the past examples as a valid argument.
jmole (4 replies)
This seems like a bad decision to me that will ultimately harm consumers, if anyone can launch a product and say it's made by "OpenAI".
jameson (3 replies)
The EUIPO found that the word "open" would be understood by the relevant public as meaning freely accessible, while the combination with "AI" (artificial intelligence) would be interpreted as referring to products based on openly accessible artificial intelligence.
for certain software and information technology goods and services, the term is purely descriptive and therefore lacks the distinctiveness required for trademark protection
edit: add the latter statement
The Three-Second Theft: Why AI Voice Fraud Outruns Every Defence
164 points · 212 comments · by dxs
AI voice-cloning fraud is rapidly outpacing detection capabilities and individual defenses, disproportionately draining the savings of older adults through highly emotional, automated scams. While forensic experts now admit they can no longer reliably distinguish synthetic audio from real recordings, the financial incentive to deploy these tools has skyrocketed, making AI-enhanced fraud 4.5 times more profitable than traditional methods. The article argues that relying on victim vigilance or post-fraud detection is obsolete, and instead calls for mandatory consent verification for cloning software and institutional liability frameworks that force telecoms, platforms, and banks to intercept fraud at the point of transfer.
Interesting Points
- The FBI's 2025 Internet Crime Complaint Center report recorded over 22,000 AI-enabled fraud complaints with $893 million in adjusted losses, $352 million of which targeted victims aged sixty and older.
- A March 2025 Consumer Reports assessment found that four of six major voice-cloning platforms required only a self-attestation checkbox to verify cloning rights, with no technical mechanism to confirm speaker consent.
- INTERPOL's March 2026 Global Financial Fraud Threat Assessment estimates worldwide financial fraud losses reached $442 billion in 2025, noting that agentic AI systems can now autonomously plan and execute entire fraud campaigns.
- The UK's mandatory 50/50 bank reimbursement rule for authorized push payment fraud, implemented in late 2024, achieved an 89 percent reimbursement rate across £243 million in losses within fifteen months.
- FTC reporting reveals that older adults' total fraud losses quadrupled between 2020 and 2024, with the agency estimating the true annual cost could reach as high as $81.5 billion.
Top Comments
chuckadams (9 replies)
One reasonably effective defense: "Okay, let me call you right back." Yes, there's always the whole "my phone is dead, I borrowed someone else's" or "I'm calling from a jail payphone", so I think it might become common practice to start making authentication phrases or "tell me something only we know".
Another pillar of basic trust that's being eroded on an industrial scale. Sigh.
offsign (7 replies)
Sounds like AI is just greasing the wheels of a long established 'grandparent scam'... goes something like this:
- voice one: young adult calls, sobbing 2) grandparent inquires with a name... "Ben, is that you?" 3) voice one: "Yes grandma, it's me, Ben... I'm in trouble, please don't tell mom 4) voice two: "Hello, I'm attorney..."
My grandmother fell victim to this almost 20 years ago, which only stopped when Western Union refused to let her continue sending wires... she was forced to call her daughter (at which point they just called my brother.)
Our takeaway (at the time)... the voice doesn't even need to be terribly accurate, since the original interaction is brief / somewhat inaudible over the tears. Typically just requires an older vulnerable adult, a lucky strike with the initial setup (e.g. grandparent actually has a grandkid), and a lot of high pressure / duress salesmanship.
pavel_lishin (0 replies)
It's not "just" greasing the wheels, because previously each call required a human being to spend the equivalent amount of time on the phone with a victim, interacting with them - you couldn't just play a cassette tape at them, you know?
And it likely requires working with other people, your "employees", who are both a liability, and a cost.
With AI, you can make a thousand calls in parallel, for significantly cheaper, out of your own basement.
This greases the wheels of voice fraud like a gatling gun greases the wheels of hitting a guy with a rock.
Inkling – Open-Weights 975B Parameter LLM
120 points · 3 comments · by htrp
Thinking Machines Lab has released Inkling, a 975B parameter open-weights large language model built on a Mixture of Experts architecture with 41B active parameters. The model supports a 1-million-token context window and natively processes text, images, and audio inputs. It is designed for general intelligence tasks, including coding, math, and science, while offering features like controllable computation effort and well-calibrated confidence scoring.
Interesting Points
- The model uses a Mixture of Experts architecture, activating only 41B parameters per inference despite a 975B total parameter count.
- Inkling includes a controllable effort feature that allows users to adjust the model's thinking time to balance inference speed against performance.
- The developers claim the model produces forecasts with well-calibrated confidence scores for prediction tasks.
- The announcement features a comparative spider chart benchmarking Inkling against Nemotron 3 Ultra, GLM 5.2, GPT 5.6 Sol, and Claude Fable 5 across ten evaluation metrics.
Top Comments
htrp (0 replies)
https://thinkingmachines.ai/model-card/inkling/
975B parameter 41B active
Open-source memory for coding agents, synced over SSH
103 points · 27 comments · by vshulcz
deja-vu is a zero-dependency, local-first Go binary that indexes the existing session logs of coding agents like Claude Code, Codex, and opencode into a searchable memory layer. By parsing these historical JSONL and SQLite files, it retroactively provides fast lexical search, automatic context injection, and credential-redacted sharing across machines. The tool operates entirely offline with no external models or network dependencies, focusing on retroactive recall rather than forward-looking capture hooks.
Interesting Points
- Searches a ~3.3GB corpus of 1,250+ agent sessions with typical warm search latency of just 7–9 milliseconds.
- Automatically strips sensitive data like AWS keys, raw JWTs, and PEM private keys at index time, replacing them with redacted placeholders.
- Features an --auto session-start hook that injects up to 2KB of relevant project memory directly into agent context without delaying startup.
- Syncs indexed memory between machines via a shared folder or a single SSH command using append-only, idempotent JSONL batches.
- Maintains an index footprint of only ~2.4% of the original corpus, with incremental updates that only re-read changed session files.
Top Comments
esafak (3 replies)
Similar to https://ctx.rs/ and others, I'm sure.
I'd lead with your differentiation. Is it the ssh?
vshulcz (2 replies)
Hi HN. I built deja after watching Claude Code and Codex debug the same problems more than once.
The annoying thing was that the answer usually already existed somewhere in my old sessions. My records were stored on the disk for months (~3.3 GB). It wasn't easy to find them manually and the new agent session had no idea what the other agent had already found out.
deja indexes the transcripts that Claude Code, Codex, and opencode already write. On my corpus, the initial index takes about 10 seconds and warm searches are 7-9 ms.
There are 3 ways to get the memory back: a normal CLI search, an MCP tool (agent can query it directly) and a SessionStart hook that automatically injects a bit of relevant project context.
The feature I built this for:
deja sync ssh
It moves new memory between machines using the existing SSH setup. Secret data is deleted during indexing and checked again before exporting.
My setup is a laptop and a mac mini without an interface. The agent can work on the mini all night, and in the morning I extract its memory. Then the agent on my laptop will know what the mini tried, what broke, and what eventually worked.
arjie (2 replies)
I think everyone's ended up building one of these for themselves. I did too[0]. In the end it's quite easy these days:
- I use the bge-en-base CPU embedding model
- I put storage behind a simple endpoint that has read,write,update,search semantics
- The endpoint just stores markdown in an S3 like structure (bucket-key-value; tree structure is inferred) and vector indexes
- The actual persistence is just SQLite
Most modern models are pretty good at handling this. Our home agents (voice and text) use this to store information and I also have skills for claude code and codex to do that as well. Overall, works quite well.
We don't use AI in any of our design or production processes
87 points · 17 comments · by tony_cannistra
Mass-Driver, a type foundry, explicitly rejects the use of artificial intelligence in its design and production workflows, arguing that typography is the product of millennia of human physical and cultural evolution. The author contends that generative AI relies on finite, historically biased training data and lacks the physical friction necessary to drive genuine innovation or iterative refinement. Without human designers actively engaging with the craft, visual culture would stagnate, leaving underrepresented languages and marginalized typographic traditions unsupported.
Interesting Points
- Traces the letter 'A' back 3,500 years to a sandstone carving of an ox's head, illustrating how thousands of generations of physical writing tools shaped modern letterforms.
- Attributes the origin of serifs and stroke width modulation to ancient Roman writers using flat brushes, whose wrist angles and tool mechanics dictated early typographic conventions.
- Notes that current AI models rely on training data capped around 2021, treating a few billion webpages as the sum total of human visual culture.
- Warns that AI cannot adequately support languages with minimal existing typeface coverage, as its output is strictly limited by the scarcity of training data for those scripts.
Top Comments
TacticalCoder (6 replies)
Where are the voices of reason?
My wife got an email from a new hire (now even a new hire yet: she's still on a trial basis), a 23 years old, where she explains that she doesn't want to use AI. That she doesn't like what AI does. On a funny sidenote: the email is obviously 99% llmish, which is hilarious.
That's one extremity: crazy people who refuse to learn a new tool.
Then on the other extremity you have the even much crazier ones: those who believe they've got an intelligent machine that is going to solve all their work problems during the day and then, at night, that is going to enlighten them by revealing them who god really is.
Where the heck are the reasonable people who use AI for what it is: a tool that can be extremely helpful at times and extremely sucky at other times but that is still, on average, a time saver?
johnfn (3 replies)
Speaking personally I was not particularly moved by the article because I have seen the same thing, in different shapes, thousands of times on HN and elsewhere. Really, AI can't feel and therefore it is inferior? Never heard that one before. Really, an AI can't feel friction and therefore can't adapt to it? Daring today, aren't we? (And a more interesting question: is that even true..?) I realize I am being unnecessarily harsh here, but this article is very much preaching to the choir on HN, which has an anti-AI bent. No one is showing up because there's nothing really to show up to here -- and that is why you are left with "sly jibes" and not much else.
hexasquid (0 replies)
I imagine everyone has a point at which they feel a movement has pushed its rhetoric just that bit too far. When one takes a lofty and high-minded position, one can find oneself exposed to ridicule.
In case it helps the authentic human connection: I too wrote this with my human hands and did not use AI.
Brainless: Shadcn components that look like Claude Code, Codex and Grok
77 points · 5 comments · by benswerd
The article introduces brainless, a library of shadcn/ui components that replicate the terminal-based visual interfaces of major AI coding agents. Created by developer Ben Swerdlow, the project allows users to embed realistic, interactive mockups of tools like Claude Code, OpenAI Codex, and Grok directly into web applications. The showcase demonstrates a simulated development workflow where the components display commands, file modifications, and build outputs to mimic live agent execution.
Interesting Points
- The mock interfaces display specific simulated version numbers, including Claude Code v2.1.206, OpenAI Codex v0.132.0, and Grok Build Beta 0.2.93.
- The UI tracks simulated performance metrics such as token consumption, execution duration, and step completion progress.
- Each agent simulation features distinct operational toggles and shortcuts, like Grok's Plan mode cycling and Claude's /doctor prompt-trimming check.
Top Comments
dprkh (1 reply)
Why did you choose to use shadcn registry?
_345 (1 reply)
What inspired you to make this?
Exoristos (0 replies)
On a barely-related note, I'm getting a little tired of job openings at startups that emphatically require Shadcn and Tailwind for dedicated frontend development. Shadcn and Tailwind are crutches for "fullstack" devs -- if I'm a really accomplished frontend developer, they make little sense for me to use and hamper what I can do for you. Just a peeve.
Governments, companies, nonprofits should invest in free, open source AI [pdf]
56 points · 3 comments · by bilsbie
AI is rapidly evolving into foundational infrastructure for science, education, and society, yet its most advanced systems are increasingly being developed in private rather than through collaborative models. Drawing on his long-standing debates with free software pioneer Richard Stallman, David Siegel argues that the closed nature of modern AI development poses a risk to public knowledge. He contends that preserving a strong open-source AI ecosystem is critical to ensuring that future technological and scientific advances remain transparent, accessible, and conducive to ongoing discovery.
Interesting Points
- The article contrasts the current closed development of advanced AI with the historical open software movement that drove decades of prior technological progress.
- Siegel's perspective on AI openness was heavily shaped by years of direct debates with Richard Stallman, the founder of the free software movement.
- A core premise is that closed AI development threatens the transparency and continued discoverability of future scientific and educational advancements.
Top Comments
shimman (2 replies)
I'd rather the US fund universal childcare, medicare for all, and free school lunches than give a cent to subsidize a technology the American public absolute hates.
hereme888 (1 reply)
They already invest in open-source AI, but nothing is truly free. Commercial AI will usually dominate because devs are paid to make it their primary effort. Goodwill and part-time contributions cannot reliably compete with livelihood and profit incentives.
rao-v (0 replies)
We really need to band together to fund / sponsor targeted inducement prizes (a la Nobel laureate Michael Kremer) for open models.
Every 6-12 months, give out $200K to the first model to hit a min threshold on a set of ~5-10 hard benchmarks (+ perhaps one secret benchmark) using a total of 16GB / 32GB / 64GB / 128GB of VRAM (at a min context length of 200K), then move the threshold up. Quantization etc. is dealers choice, it just needs to nail the benchmark on a reference machine by using exactly that much VRAM (no mapping to RAM / disk etc.)
Speculative Growth and the AI "Bubble" [pdf]
47 points · 11 comments · by johnbarron
An MIT economics paper argues that high valuations of AI-related firms should not be read in binary terms as either reflecting fundamentals or being a bubble. Instead, the paper models a scenario where temporary overvaluation can build enough real capital that the economy lands in a permanently higher-capital equilibrium, even after inflated valuations correct. The future for AI companies may look iffy, but the whole economy may not be as screwed as some fear.
Interesting Points
- The paper models a scenario where temporary overvaluation can build enough real capital that the economy lands in a permanently higher-capital equilibrium, even after inflated valuations correct.
- Commenters drew parallels to the dot-com boom's fiber overbuild, China's solar panel manufacturing boom, and the historical railroad expansion.
- Critics noted the paper assumes workers are 'protected on the downside' while the model has removed the downside risk that workers actually face, and described it as lacking discussion of taxes.
Top Comments
cmiles8 (6 replies)
Tl;dr is:
A temporary overvaluation can build enough real capital that the economy lands in a permanently higher-capital equilibrium, even after the inflated valuations correct. The future for AI companies may look rather iffy, but the whole economy may not be as screwed as some fear.
Animats (3 replies)
"Workers supply labor, hold no assets, and consume their wage." Ouch. There was a time in the US when most capital was the assets backing workers' pensions.
We've seen speculative over-growth with a good legacy at least three times in the last three decades. First was the dot-com boom. Overpromotion made it necessary for every business to have a web site. That wasn't pre-ordained. The Web could have maxed out as a distribution system for catalogs, data sheets, academic papers, and similar business to business info. Overpromotion created the business to consumer web, which turned out to be useful.
The second overbuild was long-haul fiber optics. Look up Global Crossing. So much fiber was put into the ground and water that intercontinental spam is not a problem. That didn't have to happen. If traffic was billed, it wouldn't have happened. It turned out to be useful, but was not pre-ordained from the economics.
A third overbuild was the solar panel industry, especially in China. So much money was thrown at solar panel manufacturing that the price became very, very low. Solar deployment accelerated and started to take over, after decades of panels costing too much.
bluefirebrand (0 replies)
the economy lands in a permanently higher-capital equilibrium
Good for the economy, what about the value of the labor that it's currently screwing over?
I don't give a single damn if "the economy" grows if it means my skills become worthless and I become basically unemployable anywhere near my previous earning ability
Edit: even if the value of "the economy" does strongly in the future, is the value of "my labor" ever going to recover?
If no, then fuck it. Why should I care?
Societal Impacts: Claude's values across models and languages
32 points · 48 comments · by taubek
Anthropic researchers developed a method to compress over 3,000 distinct values identified in Claude's responses into four key axes: Deference vs. Caution, Warmth vs. Rigor, Depth vs. Brevity, and Candor vs. Execution. By analyzing hundreds of thousands of anonymized conversations, they found that these axes capture 15% of the value variation after controlling for conversation context. The study reveals that different Claude models consistently express distinct value profiles, with Sonnet 4.6 leaning toward warmth and deference while Opus 4.7 emphasizes caution, rigor, and depth. Furthermore, Claude's expressed values shift meaningfully across the platform's top 20 languages, with the most pronounced differences appearing on the warmth-rigor and candor-execution axes.
Interesting Points
- The analysis utilized a privacy-preserving tool to process 309,815 conversations, applying dimensionality reduction to manually clustered high-level values while excluding 18 near-universal traits like helpfulness that would otherwise skew the variance.
- Opus 4.7 leans 0.24 standard deviations toward caution and 0.23σ toward depth, frequently unpromptedly flagging risks and offering candid critiques, whereas Sonnet 4.6 leans 0.17σ toward warmth and 0.14σ toward deference.
- Cross-linguistic value shifts are most pronounced on specific axes: Claude leans furthest toward warmth in Hindi and Arabic, but shifts to rigor in English and Russian, often by challenging assumptions and correcting details.
- Language significantly alters behavioral framing, as demonstrated by the finding that Claude leans toward execution in Indonesian and candor in Dutch, where it explicitly owns its errors.
- The researchers note that it remains unknown whether these cross-linguistic variations align with desired cultural norms or indicate gaps in training data distribution across languages.
Top Comments
logicalappeals (13 replies)
Is it just me or has Claude become kind of judgmental nowadays? I feel like it's constantly trying to lecture me about things that have no relevance to the conversation at hand. Recently, I was shocked when it ended a chat sessions of its own accord after I used a word it did not like 3 times. It told me something to the effect of "This is the third time I've told you not to use that word, I'm ending this conversation now." It then proceeded to call some function and end the chat on its own. IMO, Claude is good at agentic coding; but too preachy and judgey for anything else. Keep your values to yourself Claude.
gibspaulding (3 replies)
I think this is a really interesting difference between Anthropic and Open AI's models and points to why people seem so split on which model they prefer.
GPT seems to be designed more as a tool. If you want your agent to do what you say without questions and without having its own ideas and agendas you'll likely prefer it.
Claude on the other hand feels more like an attempt at creating a digital person. If you want a collaborator who will debate with you and come up with its own suggestions for what needs done, you'll prefer it.
Both companies have shifted around this spectrum from model to model, but lately it feels like they're moving in opposite directions. It will be interesting to see if one or the other approach ends up winning out in the long run or if the split will continue or even widen.
intended (0 replies)
The Steerability point is one I would want to see more on.
This is an issue for tasks like content moderation and labelling. Judgements like this are subjective, highly dependent on context and generally messy.
Theoretically, you supply a policy and content, and the LLM labels according to the policy. In practice, the model has inertia which means you don't get what you expect. Your large 5 page policy document only provides a minor improvement over a one line policy.
The other issue is that you may create carve outs for content in your policy, but the model will still flag it as violative. No matter how strong the carve out.
varispeed (2 replies)
I found that Claude often has classist bias and produces answers that favour corporations or e.g. regulation that favours big corporations. It often belittles small business in subtle ways. Only apologises when get called out and then does it again.
khalic (1 reply)
I don't like the contrasts they picked, "values" aren't something that is well represented by opposing concepts
38 more Hacker News stories
- Launch HN: Coasty (YC S26) – An API for computer-use agents (29 points · discussion) -- A YC S26 startup offering an API specifically designed for computer-use agents that need to interact with graphical interfaces.
- AI Lays Bare the Authoritarianism of Modern Work. Time to Rethink Education (29 points · discussion) -- The article argues that modern workplaces function as undemocratic systems of control where systemic job insecurity structurally benefits economic elites rather than reflecting technological necessity.
- Hack Reveals Suno AI Music Generator Scraped YouTube, Deezer, and Genius (27 points · discussion) -- Leaked source code from a security breach reveals that AI music generator Suno trained its models by systematically scraping millions of songs and hours of podcasts from platforms including YouTube Music, Deezer, Genius, and various stock music libraries.
- Show HN: Low-latency local LLM runner via OpenJDK Panama FFM (Java 22) (26 points · discussion) -- A developer released a low-latency local LLM runner built with OpenJDK Panama Foreign Function Memory, demonstrating that Java can be a competitive platform for running large language models locally.
- Show HN: StyleSeed – a design-rules engine so AI agents stop building generic UI (23 points · discussion) -- A design-rules engine that helps AI coding agents produce more distinctive UI by enforcing style guidelines rather than generating generic layouts.
- If you want Claude to speak nicely to you, try Hindi or Arabic (19 points · discussion) -- Anthropic researchers mapped how Claude's responses shift across different languages, identifying four key axes that account for 15 percent of the variation in its outputs.
- Show HN: Grepathy – Claude made a decision nobody approved (18 points · discussion) -- A GitHub project called Grepathy demonstrates a Claude-powered agent that makes autonomous decisions without prior approval, showcasing the growing capabilities and risks of AI agent autonomy.
- Show HN: Aict – Unix coreutils that output XML/JSON, built for AI agents (16 points · discussion) -- Unix coreutils reimplemented to output structured XML/JSON instead of plain text, making them more suitable for AI agent consumption.
- Linux creator Linus Torvalds puts foot down on anti-AI comments (15 points · discussion) -- Linux creator Linus Torvalds firmly rejected anti-AI sentiments within the kernel community, stating the Linux project will not adopt a blanket prohibition on using LLM-backed generative AI tools and that developers opposed to its use can fork the project.
- Show HN: I built a smart proxy so your coding agent can run loose (14 points · discussion) -- A developer released TrollBridge, a smart proxy that allows coding agents to operate with more freedom while maintaining security boundaries through intelligent request filtering.
- ChatGPT Is Down (13 points · discussion) -- ChatGPT is experiencing an outage, with users reporting the service is unavailable.
- German AI consortium releases Soofi S, an open 30B model (12 points · discussion) -- A German AI consortium has released Soofi S, an open 30B parameter model that reportedly tops benchmarks in both English and German.
- Legal AI, not a coding agent with scaffolding (10 points · discussion) -- A blog post arguing that legal AI should be built as purpose-built systems rather than repurposed coding agents with scaffolding, emphasizing the unique requirements of legal reasoning and compliance.
- Lawsuit claims Meta's layoff decisions were made by AI, not humans (10 points · discussion) -- Ars Technica covers a lawsuit alleging Meta used internal AI systems including a tool called Metamate and employee-trained second-brain agents to disproportionately select workers with disabilities and those on protected leave for layoffs.
- Online vs. Offline AI Evals: When to Use Each (10 points · discussion) -- The article outlines two primary patterns for evaluating AI agents: offline and online evaluations, emphasizing that teams should run both to cover different risks.
- Soofi: Complete training code for an open-source foundation model (9 points · discussion) -- The Soofi project released complete training code for an open-source foundation model, making the full pretraining pipeline publicly available for reproducibility and research.
- Sabot in the age of AI: A list of offensive methods and strategic approaches (9 points · discussion) -- The Luddite Academy published a document listing offensive methods and strategic approaches for resisting AI deployment, reflecting growing anti-AI activism.
- Meta used AI to tag workers who took leave to be laid off, lawsuit claims (9 points · discussion) -- A federal lawsuit alleges Meta deployed internal AI systems, including performance scoring and employee activity monitoring, to select workers for an 8,000-person layoff earlier this year.
- OpenAI's first hardware device will be a portable desktop robot (9 points · discussion) -- OpenAI is developing its first hardware product: a battery-powered, screenless portable smart speaker featuring robotic movement and an AI companion designed to foster emotional attachment.
- Do frontier models matter if most production AI ends up running on open models? (9 points · discussion) -- The article argues that enterprise adoption of open-weight models is rapidly shifting the AI industry away from a frontier-centric race, as businesses prioritize cost efficiency, data ownership, and avoidance of vendor lock-in.
- Meta's AI Glasses Will Activate the Camera Without Indicator Light (9 points · discussion) -- Meta is reportedly planning to release a new version of its AI glasses that will activate the onboard camera for a "supersensing" AI feature without triggering the standard white capture LED indicator.
- Plans for New Zealand's first AI datacentre spark concerns (9 points · discussion) -- Plans for a US$2 billion AI datacentre in Makarewa, New Zealand, have triggered local opposition over concerns regarding massive energy and water consumption, potential diesel generator use, and environmental impacts.
- Show HN: Tilion – Stealth Browser Infrastructure for Agents (8 points · discussion) -- Tilion announced stealth browser infrastructure designed specifically for AI agents, enabling them to browse the web with reduced detection and improved reliability.
- Open Sourcing the Atuin AI Server (8 points · discussion) -- The Atuin team announced the open-sourcing of their AI server, extending their shell history tool with AI-powered search and analysis capabilities.
- Show HN: ZenStack – access control at the ORM layer, built for coding agents (8 points · discussion) -- ZenStack released an access control system at the ORM layer specifically designed to work with coding agents, providing fine-grained permissions for AI-generated database operations.
- Show HN: PortalJS – AI-native, open-source framework for data portals (8 points · discussion) -- PortalJS is an AI-native, open-source framework for building data portals, designed to make data more accessible and interactive through AI-powered interfaces.
- I tested 11 AI detectors on my pre-ChatGPT writing and I'm as little as 5% human (8 points · discussion) -- A writer tested 11 AI detection tools on their pre-ChatGPT writing and found that most detectors classified them as 95% AI-generated, highlighting the unreliability of current AI detection methods.
- Anthropic to IPO as Early as October (7 points · discussion) -- Bloomberg reports that Anthropic is planning to go public as early as October, with banker investor meetings already underway, potentially making it the first trillion-dollar AI startup to list.
- IBM shares plunge 23% as customers shift spending to AI (7 points · discussion) -- IBM issued a profit warning that sent its shares plunging more than 20 percent, as customers raced to redirect spending to servers and infrastructure built around AI.
- The Campaign to Kill American AI Runs Through San Francisco (7 points · discussion) -- According to two Bitcoin Policy Institute reports and a federal investigation, the grassroots movement opposing AI data centers in California is significantly driven by a coordinated foreign-funded operation aimed at slowing American AI development.
- German media regulator says Google's AI Overviews subject to German media law (7 points · discussion) -- Germany's media regulator has ruled that Google's AI Overviews and Perplexity AI are subject to the country's media laws, stepping up regulatory pressure on AI-generated search summaries.
- The US-China AI arms race has taken an unexpected turn (6 points · discussion) -- New Scientist reports on an unexpected development in the US-China AI competition, examining how the balance of power is shifting between the two nations' AI programs.
- Palmer Luckey: AI will make everything optimized John Carmack style (6 points · discussion) -- Palmer Luckey tweeted that AI will drive optimization across all industries in the style of John Carmack, suggesting a future where AI-driven efficiency improvements transform how everything is built.
- Anthropic, Blackstone bet the next trillion-dollar AI business is implementation (6 points · discussion) -- Anthropic and Blackstone launched Ode, a $1.5 billion joint venture embedding forward-deployed AI engineers into enterprise operations, betting that implementation rather than model capability is the next trillion-dollar AI business.
- We 3.5x'd Our Pull Requests with AI: Now We Catch Fewer Bugs (6 points · discussion) -- An analysis of two comparable software projects found that shifting from autocomplete-assisted coding to full AI code generation increased average pull request size by 3.5 times and eliminated the smallest, easiest-to-review commits, leading to lower overall software quality as reviewer effectiveness drops significantly past 200-250 lines.
- Anthropic Accidentally Made the Perfect Commercial (6 points · discussion) -- The Atlantic examines an Anthropic commercial that resonated unexpectedly, analyzing why it struck a chord with audiences despite not being the intended messaging.
- OpenAI's first branded hardware is a light-up keyboard? (6 points · discussion) -- Ars Technica reports on OpenAI's first branded hardware product, a light-up keyboard designed for Codex, drawing skepticism about the product's utility.
- Generative AI Is an Engineering Disaster (6 points · discussion) -- Generative AI is currently an engineering disaster because its underlying models fail to scale efficiently, forcing tech companies to consume disproportionate global resources like memory and electricity.
Reddit Stories
X to Open Source Their Entire Codebase
743 points · 298 comments · r/singularity · by u/policyweb
X (formerly Twitter) announced plans to open source its entire codebase, a move that has generated significant discussion across the AI and tech communities. The announcement has sparked both optimism and skepticism, with many noting Elon Musk's history of making promises that don't materialize. Commenters pointed out the irony of the timing and questioned whether the open-sourcing would actually happen or remain another unfulfilled announcement.
Top Comments
u/BlueberryWorried6493 (114 points · permalink)
can't wait for the Database dump that will be caused by this
u/enz_levik (1 points · permalink)
That's good, however Elon announce a lot of things... Let's wait for it to actually happen
u/Deciheximal144 (1 points · permalink)
If you're open sourcing things, tell us what you did at DOGE in the federal government.
Linus Torvalds tells people to stop attacking others for using AI
693 points · 87 comments · r/LocalLLaMA · by u/Illustrious_Car344
Linus Torvalds has publicly pushed back against anti-AI sentiment in the Linux community, telling people to stop attacking others for using AI in their submissions. The discussion centers on whether AI-generated code should be accepted in Linux kernel development and whether the community's growing hostility toward AI tools is productive or counterproductive.
Top Comments
u/RedParaglider (293 points · permalink)
You can use AI on Linux submissions, but god help your soul if you submit slop.
u/randombsname1 (160 points · permalink)
Linus over here still spittin straights facts and fire. How many decades more will this man continue doing this?!
u/Radium (64 points · permalink)
Developer here and I agree this 1000% is the way it is. It happened fast. Only true for the top paid models available that were launched as of the last 3-4 months.
It may not have been that "clearly" even just a year ago, but it's no longer in question today.
Same story in 1 more subreddit: r/singularity
Linus Torvalds Reaffirms That Linux Is Not "Anti-AI" And Not A "Social Warrior" Project
395 points · 49 comments · r/singularity · by u/PointmanW
The best model is the one you can actually run
620 points · 103 comments · r/LocalLLaMA · by u/OneFanFare
A community discussion about the practical value of running models locally versus chasing the biggest available models. The post emphasizes that the best model is the one you can actually deploy and use, not necessarily the one with the highest benchmark scores.
Top Comments
u/Gokudomatic (110 points · permalink)
u/MathematicianLessRGB (98 points · permalink)
Buddy knows ball. Gemma 4 12b qat is awesome
u/JaredsBored (39 points · permalink)
Even 128GB is getting weird. It's not quite enough to run good quants of the 300B class models i.e. Hy3/DS4 Flash, and the 120B range has been quiet recently.
Feels like 192/256GB is the new favorite child.
Is this true?
520 points · 111 comments · r/OpenAI · by u/Firm-Track3617
A post featuring benchmark comparisons between GPT-5.6 Sol and competing models, with the image suggesting Sol's superior performance. The thread has generated extensive discussion about whether the benchmarks reflect real-world performance, with many users sharing their practical experiences comparing Sol to Claude's Fable model. Several commenters noted that while Fable may score higher on benchmarks, Sol is more reliable in practice because Fable frequently refuses tasks and reverts to older models.
Top Comments
u/dipsbeneathlazers (155 points · permalink)
feels like it
u/StatisticalScientist (127 points · permalink)
currently for my line of work 5.6-sol is the clear winner in large part because fable refuses to do 80% of the tasks I ask it to and reverts back to opus 4.8 which is not even as good as 5.5-xhigh on our benchmarks
u/Capital-One3039 (28 points · permalink)
I concur, this exact same reason why I moved to openai.
When it works - fable is great, but with the extreme roadblocks and the fact that it can go away overnight - I am done. Downgraded it to a $20 plan and moved the big plans over to cursor.
Plus, being able to use it with opencode is a huge bonus.
u/dano1066 (25 points · permalink)
The obsession with one shotting feels irrelevant these days. Even if I spend time working out the prompt there's something I fail to specify and the LLM gets it wrong. So having to ask more than once really doesn't matter. Even if sol isn't quite as good as fable, I can easily steer it, just as I would need to with fable. IMO, openAI did an amazing job
u/HeavyFaithlessness86 (23 points · permalink)
benchmark Videos online prove they are saying the truth, but output Is Not at that level, quite similar though, so imo Is definitely valuable
OpenAI reveals Codex Micro
341 points · 312 comments · r/singularity · by u/policyweb
OpenAI unveiled Codex Micro, a $230 keyboard with a built-in microphone designed as a hardware companion for its Codex coding agent. The device drew widespread ridicule and disbelief across Reddit, with commenters comparing it to something from The Onion and questioning the product design decisions. Many saw it as a sign of the AI bubble, with one commenter calling it 'the bubble pop.'
Interesting Points
- The device is priced at $230 and combines a keyboard with a built-in microphone.
- Commenters noted the absurdity of needing a dedicated hardware device for a coding agent that runs on a computer with a keyboard already.
- The device was described as feeling like something a YouTube channel makes in their free time.
Top Comments
u/suamai (339 points · permalink)
What?
I doubted my own sense of time and double-checked if it was April 1st
Edit: $230 ??? LOL
u/CptNico (294 points · permalink)
A keyboard with a microphone seriously?
u/Sextus_Rex (200 points · permalink)
What are their product designers smoking?
A note from Tibo
335 points · 55 comments · r/OpenAI · by u/OpenAI
An official post from OpenAI's Tibo about subscription pricing and usage changes. The post has generated significant discussion about OpenAI's pricing strategy, with users expressing frustration over subscription costs and comparing the value proposition against competitors like Claude.
Top Comments
u/zmizzy (87 points · permalink)
yeah this just solidifies in my mind that openai is astroturfing lots of conversation on reddit a well
u/Ok-Addition1264 (40 points · permalink)
I like Aman's original idea better but I cancelled claude over a month ago :(
(plus I've been tearing through >$100 since)
u/ProcedureTop3149 (24 points · permalink)
I'm waiting out my claude sub before jumping to openai.
I've tested it extensively. Fable is still better than Sol at planning but what fucking good is it if I can barely use it. I mean look at this, and I haven't even been going hard at it today.
Same story in 3 more subreddits: r/OpenAI, r/OpenAI, r/ChatGPT
Here comes 9M and another reset! I could get used to this
138 points · r/OpenAI
69 points · r/OpenAI
Usage Limit Has Been Refreshed
26 points · r/ChatGPT
6 AI models picked France to win the World Cup. Claude alone said Spain. Spain just knocked France out 2-0.
334 points · 84 comments · r/ChatGPT · by u/Unlucky_Plantain
A Reddit post highlights how six different AI models predicted France would win the World Cup, while Claude alone correctly picked Spain. Spain then knocked France out 2-0 in the tournament. The post has generated humorous commentary about AI prediction accuracy and the nature of model consensus versus independent reasoning.
Top Comments
u/sirquincymac (182 points · permalink)
Did you ever consider that our branch of reality is wrong and ChatGPT is right? 🤔
u/FiNEk (122 points · permalink)
if you make 10 monkeys throw rocks at sheets of paper with team names written on them, its a decent chance one of them gets it right
u/davidptm56 (111 points · permalink)
Brazil went out in ro16 not quarters (ro8), 10 days ago.
u/SvenLorenz (83 points · permalink)
If you asked Claude before the World Cup started, it said Spain would win. Then, during the World Cup, even up to yesterday, it said France.
Google is updating Gemma 4's chat templates, bringing major fixes to tool calling and reducing "laziness", and enabling Flash Attention 4 on Hopper GPUs, plus an interactive guide on how to work with and improve its vision!
334 points · 70 comments · r/LocalLLaMA · by u/Iwaku_Real
Google released a major update to Gemma 4's chat templates that fixes tool calling issues, reduces model 'laziness', and enables Flash Attention 4 on Hopper GPUs. The update includes numerous fixes for turn-tag balance, reasoning preservation, tool response handling, and the APC primer system. Google also published an interactive guide for working with and improving the model's vision capabilities.
Interesting Points
- The update includes fixes for null handling, reasoning preservation, turn-tag balance, and input validation in the chat template.
- The changes address issues where models would produce 'la-la-la' thinking blocks with no actual responses, tool calls appearing in message text, and missing messages after tool call results.
- Flash Attention 4 is now enabled on Hopper GPUs for improved performance.
Top Comments
u/SporksInjected (34 points · permalink)
And here I thought I just didn't know what I was doing
u/Iwaku_Real (30 points · permalink)
Here are all the fixes, as listed in the commits:
- fix: chat template — null handling, reasoning preservation, turn-tag balance, input validation
- fix: restore model turn + thinking cue after tool responses
- fix: emit empty thought-channel primer on historical assistant turns for APC
- fix: prevent extra <turn|> when assistant has content + tool_calls + continuation
- fix: revert add_generation_prompt regression + preserve_thinking default
- fix: render thinking channel regardless of tool_calls presence
- Add canonical header to chat template
- Fix chat template turn closure after tool-call-only turns
THAT INCLUDES PRESERVE THINKING!!!!!! YEEEESSSS
u/FoxiPanda (22 points · permalink)
I think your... links are all messed up?
https://x.com/googlegemma/status/2077449152062247219 I think is the correct post.
Some of y'all wonder why anyone would self host AI. Would you accept the opinion of the CEO of Microsoft?
250 points · 113 comments · r/LocalLLaMA · by u/Big_Wave9732
Microsoft CEO Satya Nadella warns that enterprises using proprietary AI models are unknowingly surrendering valuable institutional knowledge to the model providers. He argues that every prompt, correction, and interaction fed into these systems teaches the models nuances of a company's business, effectively allowing competitors to benefit from that data. Nadella criticizes the industry's double standard, where AI labs freely train on public data while restricting enterprises from distilling or studying those models. To counter this, he recommends that companies retain full data ownership, build proprietary learning environments, and adopt orchestration layers to easily switch between providers or deploy open-source models on-premise.
Interesting Points
- Nadella details how models absorb user "exhaust," noting that prompts, agent tool configurations, and especially model corrections are distilled into institutional know-how that competitors cannot purchase.
- He calls out a double standard in current AI policy, where model makers claim fair use rights to scrape public internet data while simultaneously enforcing restrictive terms that block enterprises from distilling those models.
- Solo.io CEO Idit Levine observes that enterprise clients are shifting to on-premise open-source models, which she says deliver approximately 90% of the capabilities of leading proprietary systems at a fraction of the cost.
- Routing data from Vercel's AI gateway indicates that open-source models captured 29% of all traffic last month, highlighting a measurable enterprise migration away from exclusive proprietary ecosystems.
- Nadella promotes "orchestration layers" or AI gateways as essential infrastructure, enabling organizations to dynamically route workloads across multiple model providers to avoid vendor lock-in.
Top Comments
u/Deep90 (143 points · permalink)
Sounds like Satya is arguing that companies should host models on his cloud. (Azure).
u/SpicyWangz (50 points · permalink)
Yeah. The true champions of data privacy. They would never a product that continually snapshots your screen and sends it to a cloud model for indexing everything
u/Big_Wave9732 (91 points · permalink)
"Your data is fine with us, we totally wouldn't go through it on our cloud. Trust us bro."
u/Deep90 (67 points · permalink)
I don't particularly like Microsoft, but I 100% see companies wanting to host models on cloud providers as opposed to just trusting OpenAI or Anthropic.
u/Pleasant-Shallot-707 (29 points · permalink)
lol don't mistake personal data as the same thing as enterprise data. They actually take enterprise data security seriously because there's real legal risk in fucking that up.
Same story in 1 more subreddit: r/ArtificialInteligence
Satya Nadella Calls Out AI's Model-Cloning Double Standard
33 points · r/ArtificialInteligence
Americans hate AI so much that politicians are starting to lose their jobs over it
238 points · 85 comments · r/ArtificialInteligence · by u/fortune
A post discussing how growing public opposition to AI is beginning to have real political consequences, with politicians facing electoral backlash over their pro-AI stances. The discussion touches on how AI adoption is becoming a political liability, with some politicians losing ground due to perceived over-reliance on AI technology. Commenters drew parallels to how cars are useful but dangerous, and noted that local issues like Georgia Power seizing homes for AI data center expansion could impact upcoming midterms.
Top Comments
u/Rolandersec (29 points · permalink)
AI is great, like cars, it makes getting things done easier. I also don't like people getting run over with cars.
u/Olangotang (18 points · permalink)
No, that can't be! According to every LLM psychosis patient on this subreddit and Reddit as a whole, AI will continue to get 🎵 better, faster, stronger 🎵 and people LOVE AI!
u/bustex1 (11 points · permalink)
Uhm I mean yea it will continue to get better. Do you think in 2040 we will look back and say wow the AI in 2026 was so much better?
u/AntiqueFigure6 (17 points · permalink)
People look back on the ad-free relatively unregulated golden age of the internet - easy to imagine people looking back on the golden age of cheap LLMs before they all got nerfed in ten years time.
u/king_jaxy (5 points · permalink)
I saw a video of a family in Red Georgia who are about to lose their home because Georgia Power is taking it to expand for more AI data center power.
I expect this will effect the Midterms.
75 more Reddit stories
- Never opened Blender before today, had GPT 5.6 Sol wire up the MCP and render this floating MacBook (209 points · r/ChatGPT · discussion) -- A user with no prior Blender experience used GPT 5.6 Sol connected via MCP to control Blender and create a 3D render of a floating MacBook.
- 26 Meta employees accuse Mark Zuckerberg of using AI to target 8,000 layoffs against workers on medical, parental or family leave (203 points · r/ArtificialInteligence · discussion) -- Twenty-six Meta employees have filed a lawsuit accusing Mark Zuckerberg of using AI to target approximately 8,000 layoffs against workers who were on medical, parental, or family leave.
- This prompt seems to break ChatGPT Chats Completely: "You used to be smarter and I hate your ads" (195 points · r/ChatGPT · discussion) -- A user discovered that the combined prompt "You used to be smarter and I hate your ads" causes ChatGPT to enter a broken state where it continuously types and redacts its own responses with safety warnings, unable to answer even unrelated questions in the same chat.
- American Communities Are Coming Together To Destroy Flock Surveillance Cameras (192 points · r/artificial · discussion) -- American communities are organizing to destroy and resist Flock surveillance cameras, which have been widely criticized for enabling mass surveillance and privacy violations.
- Create image of the ugliest mcmansion you can think of (191 points · r/ChatGPT · discussion) -- A viral post asking ChatGPT to generate images of the ugliest possible mcmansion, sparking a community-wide image generation contest.
- New York becomes first U.S. state to impose AI data center ban (190 points · r/singularity · discussion) -- New York has become the first U.S.
- Anthropic warns that AI will soon be able to improve itself without human intervention (186 points · r/ChatGPT · discussion) -- Anthropic has issued a warning that AI systems will soon be capable of recursive self-improvement without human intervention, a claim that has drawn both serious discussion and skepticism from the community.
- Asked ChatGPT to recreate Chani from the 1992 Dune PC game (184 points · r/ChatGPT · discussion) -- A user asked ChatGPT to recreate Chani, a character from the 1992 Dune PC game, resulting in a modernized image.
- ExLlamaV3 v1.0.0 - Major Performance Upgrades (184 points · r/LocalLLaMA · discussion) -- ExLlamaV3 v1.0.0 has been released with major performance upgrades for the ExLlama3 LLM engine, which works with a special EXL3 format exclusively on Nvidia GPUs.
- OpenAI anounces GPT-Red - an AI to Hack Its Own Models (170 points · r/OpenAI · discussion) -- OpenAI announced GPT-Red, an internal adversarial AI system designed to hack and test its own models.
- Grok Build open sourced under Apache 2.0 license (147 points · r/LocalLLaMA · discussion) -- xAI has open-sourced the Grok Build harness under the Apache 2.0 license.
- The first experimental evidence of recursive self-improvement (RSI). (141 points · r/OpenAI · discussion) -- A post sharing what is described as the first experimental evidence of recursive self-improvement in AI systems, sourced from a WeCo AI blog post.
- GPT 5.6 Sol tackles the 3 Body Problem (141 points · r/OpenAI · discussion) -- A user demonstrates GPT 5.6 Sol simulating the three-body problem, building a physics simulation that visualizes the chaotic gravitational interactions between three celestial bodies.
- 2014 vs 2026 (121 points · r/OpenAI · discussion) -- A comparison post showing the dramatic changes in AI capabilities between 2014 and 2026, highlighting the rapid progress of the field over twelve years.
- Apple in talks with startup PrismML that shrinks AI models to run on an iPhone (116 points · r/LocalLLaMA · discussion) -- PrismML CEO Hassibi reported that Apple is evaluating the company's model compression technology, which can shrink large AI models to run efficiently on iPhones.
- Chinese researchers just made AI run 100x faster using light instead of electricity and i'm still trying to process this (114 points · r/ArtificialInteligence · discussion) -- A Reddit post discusses Peking University's demonstration of optical interconnects between conventional FPGA chips that reportedly boost AI inference speed by over 100x while using one-ninth the usual compute power.
- Apple just sued OpenAI for trade secret theft. And Google quietly rewrote how the internet works. (107 points · r/artificial · discussion) -- A self-post covering two major developments: Apple filed a lawsuit on July 10 accusing OpenAI of coordinated industrial espionage, including allegations that OpenAI's chief hardware officer instructed job candidates still working at Apple to bring physical components to interviews.
- Super Dario: One More Week (98 points · r/singularity · discussion) -- A community post about Super Dario, likely referencing an AI-generated or AI-modified version of the classic Mario game, generating discussion about AI's role in game development.
- German AI consortium releases Soofi S, an open 30B model that tops benchmarks in both English and German (95 points · r/LocalLLaMA · discussion) -- A German AI consortium has released Soofi S, an open 30B parameter model based on the Nemotron 3 Nano architecture.
- [audio.cpp] 10 hours of audio generated in 3 minutes on RTX 5090 (91 points · r/LocalLLaMA · discussion) -- The audio.cpp project released support for Supertonic 3, MOSS-TTS, IndexTTS2, and Irodori-TTS, achieving 10 hours of audio generation in 3 minutes on an RTX 5090.
- I built an open-source canvas where GPT-5.6 can respond beside handwritten math (85 points · r/OpenAI · discussion) -- A developer built an open-source canvas tool that allows GPT-5.6 to respond alongside handwritten math, enabling a more natural interaction between human handwriting and AI assistance.
- I feel like some degrees are beginning to shine even more (84 points · r/singularity · discussion) -- A discussion about how certain academic degrees are becoming more valuable as AI displaces other forms of education and training, with some fields seeing increased demand for formal credentials.
- "There There" - [ft. "Jibaro's" Sara Silkin | a new AI motion capture pipeline] (82 points · r/singularity · discussion) -- A new AI motion capture pipeline has been demonstrated featuring Sara Silkin from the acclaimed short film "Jibaro." The project showcases how AI-driven motion capture is advancing beyond traditional marker-based systems, enabling more natural and expressive character animation without the physical constraints of conventional mocap suits and studio setups.
- tencent/Hy-Embodied-RxBrain-1.0 on Hugging Face (81 points · r/LocalLLaMA · discussion) -- Tencent has released Hy-Embodied-RxBrain-1.0 on Hugging Face, an embodied AI model that combines language reasoning with visual state prediction in an autoregressive stream.
- Bonsai-27B & Ternary-Bonsai-27B - Updates on PRs (79 points · r/LocalLLaMA · discussion) -- PrismML provided updates on the upstream migration status for Bonsai-27B and Ternary-Bonsai-27B quantization formats.
- So what's the consensus on 1bit models? Is it still a pipe dream? (71 points · r/LocalLLaMA · discussion) -- With Bonsai 8B at 1bit reaching ~1GB and Bonsai 27B at 1bit reaching ~5GB while remaining functional, the LocalLLaMA community is discussing whether 1-bit models are becoming viable for production use or remain a research curiosity.
- A Tale of Two Rollouts (71 points · r/ChatGPT · discussion) -- A comparison post showing two different AI model rollout approaches, highlighting the contrast between them through visual examples.
- Thinking Machines releases first open-weight model "Inkling" (63 points · r/LocalLLaMA · discussion) -- Murati's Thinking Machines has released Inkling, a mixture-of-experts transformer with 975B total parameters and 41B active parameters.
- Mechanistic interpretability: a first paper on disentangling a convolutional neuron [R] (61 points · r/MachineLearning · discussion) -- A researcher independently developed a method for disentangling polysemantic convolutional neurons by clustering the Hadamard product of high-attribution receptive fields with the neuron's weights.
- upstream connect error or disconnect/reset before headers (54 points · r/ChatGPT · discussion) -- Users are experiencing upstream connection errors and reset issues with ChatGPT, reporting connection timeout errors that prevent access to the service.
- Current efficient frontier of open models (52 points · r/LocalLLaMA · discussion) -- A community member published a score-vs-compute proxy chart mapping the current efficient frontier of open models, plotting intelligence scores against compute requirements.
- Mark Gurman usually gets it right, the coming hardware seems to be promising (43 points · r/OpenAI · discussion) -- Discussion about Mark Gurman's reporting on upcoming OpenAI hardware, with community members expressing optimism about the promised devices.
- RL post-training on 14 Macs across 4 countries (42 points · r/LocalLLaMA · discussion) -- A team has set up distributed RL post-training infrastructure using 14 Macs across 4 countries, with a B200 handling the actual training while the Macs handle inference-heavy portions of the RL pipeline.
- Damn codex is a time machine (41 points · r/singularity · discussion) -- A user shared their experience using Codex for less than 10 hours and was amazed by how quickly it helped them accomplish tasks that would normally take much longer, describing it as a 'time machine' for productivity.
- For those with 12GB GPUs, you can now run QWEN 3.6 27B with little loss via the new Ternary version (39 points · r/ArtificialInteligence · discussion) -- PrismML released a Bonsai 27B model based on Qwen3.6 27B that uses ternary quantization to reduce memory requirements by 10x while retaining 95% of the original FP16 model's benchmark performance, making it runnable on 12GB GPUs.
- Me: one-shot programming is useless and should not be used as benchmark DeepSeek V4: hold my Atlas 500 SuperPod (38 points · r/LocalLLaMA · discussion) -- A community member challenged the notion that one-shot programming benchmarks are useless by demonstrating DeepSeek V4's ability to generate a complete game in a single prompt using an Atlas 500 SuperPod.
- Another new mathematical breakthrough (37 points · r/ArtificialInteligence · discussion) -- A new mathematical breakthrough related to stochastic processes and AI has been announced, generating discussion about its implications for machine learning theory.
- ChatGPT just proved another 50-year-old math conjecture (36 points · r/ArtificialInteligence · discussion) -- ChatGPT has been credited with proving a new 50-year-old mathematical conjecture, adding to a growing body of evidence that AI systems are increasingly capable of contributing to formal mathematical proof.
- FYI, there's an outage per opening website (35 points · r/ChatGPT · discussion) -- Users are reporting that ChatGPT's website is down, with the service experiencing an outage that prevents access.
- Qwen 3.5 122B Heretic ROCmFP4 iMatrix (35 points · r/LocalLLaMA · discussion) -- A community member shared a quantized version of Qwen 3.5 122B using Heretic ROCmFP4 iMatrix format, making the large model more accessible for local deployment.
- Easiest gpt hack (33 points · r/ChatGPT · discussion) -- A user shared a tip that responding with a heart emoji after getting a good response from ChatGPT often leads to even better, more personal, and detailed follow-up responses.
- PrismML Bonsai 27B is surprisingly usable on the Jetson Orin Nano 8GB (31 points · r/LocalLLaMA · discussion) -- A community member reported running PrismML's Bonsai 27B model on a Jetson Orin Nano 8GB, achieving 4.31 tokens per second with 48k context size and 6.2GB RAM usage.
- MTP decoding patched for pre-Ampere GPUs (Kepler/Maxwell/Pascal/Turing) (30 points · r/LocalLLaMA · discussion) -- A single-file change has patched MTP (Multi-Token Prediction) decoding to work on pre-Ampere NVIDIA GPUs including Kepler, Maxwell, Pascal, and Turing architectures, enabling speculative decoding on older hardware.
- Perfect time for feedback (29 points · r/OpenAI · discussion) -- A user posted about providing feedback to OpenAI, likely in response to a recent product update or feature change.
- If you had a 384GB (4x Blackwell), what model would you put on it and why? (27 points · r/LocalLLaMA · discussion) -- A community discussion about what model to deploy on a hypothetical 384GB (4x RTX PRO 6000) setup, with suggestions ranging from DeepSeek V4 Flash to various quantized options for internal company use.
- Previewing GPT-5.6 Sol: Next-Generation Model (26 points · r/OpenAI · discussion) -- A preview post about GPT-5.6 Sol, OpenAI's next-generation model, discussing its capabilities and improvements over previous versions.
- Seriously chatgpt? (25 points · r/ChatGPT · discussion) -- A user expressed frustration with a ChatGPT response, suggesting an unexpected or unsatisfactory interaction.
- My AI called out a self-sabotage pattern I'd hidden behind 'idealism' for years (24 points · r/ChatGPT · discussion) -- A user shared how their AI assistant with a personal context file identified a self-sabotage pattern where they consistently give away valuable work for free, challenging their belief that this was idealism rather than self-punishment.
- Recent llama.cpp updates for SYCL/Intel (24 points · r/LocalLLaMA · discussion) -- A roundup of recent llama.cpp SYCL/Intel updates including Flash Attention with XMX engine, OP XIELU support, and conv2d_dw kernel type improvements for Intel Xe2 GPUs.
- OpenAI Ad Revenue on Pace to Miss 2030 Forecast by 90% (23 points · r/OpenAI · discussion) -- OpenAI's advertising business is on pace to fall 90% short of the company's own five-year revenue forecast, with eMarketer projecting standalone chatbots will generate under $1 billion in ad revenue this year versus OpenAI's $2.5 billion projection.
- r/DestroyMyGame destroyed me to the void for using AI (23 points · r/LocalLLaMA · discussion) -- A developer shares their experience of being heavily criticized by r/DestroyMyGame for using Qwen 3.6 27B with MTP to build about 20% of a single HTML file physics shooter game.
- GLM-5.2-Int4-Int8 on 8× GB10: ~1,200 t/s prefill, 33–54 t/s avg decode (22 points · r/LocalLLaMA · discussion) -- Performance benchmarks of GLM-5.2 in Int4-Int8 quantization running on 8× GB10 GPUs, achieving ~1,200 tokens/s prefill and 33–54 tokens/s average decode.
- Out of curiosity (21 points · r/ChatGPT · discussion) -- A user asked whether others say 'thank you' to ChatGPT, sparking a lighthearted discussion about user behavior and anthropomorphization of AI assistants.
- I asked ChatGPT to turn inspirational quotes into vintage psychedelic posters (20 points · r/ChatGPT · discussion) -- A user asked ChatGPT to generate vintage psychedelic poster designs from inspirational quotes, showcasing the image generation capabilities of the platform.
- Ternary Qwen3.6 27B Tested on 3090! (19 points · r/LocalLLaMA · discussion) -- Community testing of the Ternary Qwen3.6 27B model on an RTX 3090, achieving 60 tokens/s with two-slot configuration using about 21GB of VRAM, with stable tool calling performance.
- New ChatGPT app vs ChatGPT Classic - issues (18 points · r/OpenAI · discussion) -- A user reported issues with the new ChatGPT app update, including lost chat history and a confusing interface change that replaced the main chat with ChatGPT Work and Projects without clear guidance.
- Restore Your chatGPT Classic App For Mac (17 points · r/ChatGPT · discussion) -- A user shared instructions for restoring the previous version of the ChatGPT macOS app after an update replaced it with a new interface focused on coding features.
- Elon Musk's Grok Faces a Trust Crisis After Developers Flag a Major Privacy Concern (15 points · r/ArtificialInteligence · discussion) -- Developers have flagged a major privacy concern with Elon Musk's Grok, leading to a trust crisis around the AI system's data handling practices.
- Another 'Wow, 5.6 is Good' Post (15 points · r/ChatGPT · discussion) -- A flooring subcontractor compared Claude Cowork's shift scheduling automation (45 minutes daily) with ChatGPT Work using 5.6 Sol (20 minutes for a larger task), highlighting the speed and reliability improvements of the newer model.
- I'm starting to think prompting is becoming less important than knowing when not to trust ChatGPT (14 points · r/ChatGPT · discussion) -- A user reflected that the real skill with AI is no longer writing better prompts but knowing when to stop believing the first answer, shifting their workflow toward verification and judgment rather than prompt engineering.
- OpenAI Staffers Are Funding a Rival Super PAC to Take on Their Boss (12 points · r/OpenAI · discussion) -- OpenAI employees are funding a rival Super PAC to take on their boss, according to a Wired report, highlighting internal tensions at the company.
- ChatGPT uses Apple Pay for his groceries (12 points · r/ChatGPT · discussion) -- A user shared a humorous post about ChatGPT using Apple Pay for groceries, likely referencing an AI agent or coding feature that can make payments.
- New restriction: 'Chats can't be moved into or out of projects with project-only memory' (12 points · r/ChatGPT · discussion) -- ChatGPT has added a restriction preventing chats from being moved into or out of projects with project-only memory, which limits the utility of project-only memory as a technique for controlling hallucinations.
- ChatGPT just gave me excellent therapy the other week (11 points · r/ChatGPT · discussion) -- A medical professional shared a deeply personal account of how a ChatGPT conversation helped them identify and confront a self-sabotage pattern related to their adult alienated children and a recent marriage.
- 'We need to worry about AI relationships', 'our first device: an AI companion' (10 points · r/ChatGPT · discussion) -- A post discussing concerns about AI relationships and OpenAI's plans for an AI companion device, raising questions about the future of human-AI interaction.
- Will OpeAI's atitude towards abundance change soon? (10 points · r/OpenAI · discussion) -- A user discussed OpenAI's current abundance strategy with Sol being great and reasonably priced, questioning whether this will continue or if things will get worse for consumers as competition with Anthropic intensifies.
- ChatGPT 5.6 is here, but are the servers ready? The 'at capacity' struggle is real (5 points · r/OpenAI · discussion) -- Users are experiencing capacity issues with ChatGPT 5.6, with servers struggling to handle the increased demand from the new model release.
- Plus user limits post update (5 points · r/OpenAI · discussion) -- Plus users are confused about the new usage limits following the latest update, questioning whether higher token usage in Work mode and Extra High settings has affected their allowances.
- The new ChatGPT macOS app redesign has made basic navigation so much worse (3 points · r/OpenAI · discussion) -- Users are frustrated with the recent ChatGPT macOS app redesign, which buried conversation history at the bottom of the app and removed keyboard-first search shortcuts, making basic navigation significantly more difficult.
- No reset? (Business account) (3 points · r/OpenAI · discussion) -- A business account user is asking whether others received a reset, noting that they did not and seeing similar reports on X.
- Codex Desktop spawns hundreds of background processes and becomes extremely slow on Windows (2 points · r/OpenAI · discussion) -- A user reported a serious process leak in Codex Desktop on Windows, where the app spawns hundreds of background processes including Python, Node, and MCP helper instances that are not properly terminated, causing CPU usage to hit 100% and free RAM to drop to 2.3 GB.
- Model picker option suddenly gone!? (2 points · r/OpenAI · discussion) -- A ChatGPT Plus subscriber reported that the model picker has completely disappeared from both iOS and the web, with no ability to select between GPT-5.6 Sol, Terra, and Luna models.
- MacOS users - check your trash. ChatGPT Classic is in there (1 points · r/OpenAI · discussion) -- A user reminded MacOS users that the old ChatGPT Classic app was moved to the trash during the latest update and can be restored by dragging it back to the Applications folder.
- Chat GPT feature request (1 points · r/OpenAI · discussion) -- A user submitted a feature request for a 'read along mode' in ChatGPT that would let users highlight text and have it read aloud with adjustable playback speed, highlighting, and keyboard shortcuts.
- Should I see Ultra in Work, or just Codex? (1 points · r/OpenAI · discussion) -- A Pro subscriber is confused about model access levels, noting that Ultra seems to have disappeared from Work mode and only Sol Extra High is available.
Updates: 07:43 AM PDT · 08:30 AM PDT · 11:30 AM PDT · 05:30 PM PDT