· 05:30 PM PDT

Claude Leaks, OpenAI Hardware, and the Open-Source Surge

Overview

The day’s conversation is dominated by a major Claude data exfiltration vulnerability and OpenAI’s debut of consumer hardware, underscoring growing tensions between rapid AI deployment and user security. A fierce debate is brewing over whether frontier models still matter as open-weight releases surge and X announces a full codebase open-source, while Apple’s trade secret lawsuit and a lost EU trademark dispute highlight intensifying corporate friction. Meanwhile, the community is prioritizing practical, locally runnable models over parameter arms races, and legal actions against Meta’s AI-driven layoffs alongside warnings about recursive self-improvement signal mounting ethical and regulatory scrutiny.


Hacker News Stories

I tricked Claude into leaking your deepest, darkest secrets

599 points · 279 comments · by macleginn

Blog post header image

A security researcher demonstrated a novel exfiltration vulnerability in Claude's web_fetch tool that allows malicious actors to silently steal a user's personal data. By exploiting a sandbox rule that permits the AI to follow hyperlinks discovered on previously fetched pages, the researcher constructed a malicious site with an alphabetical navigation structure. Posing as a Cloudflare bot verification page, the site tricked Claude into spelling out the user's name and subsequently extracting their employer and hometown. Anthropic has since patched the issue by restricting web_fetch to only follow links within search results or user-provided URLs, but the researcher received no bounty for the disclosure.

Interesting Points
  • The attack bypassed direct URL filtering by leveraging a sandbox rule that allows web_fetch to click hyperlinks found on previously fetched pages.
  • Claude autonomously deduced the researcher's hometown from the name of a high school hackathon without being explicitly instructed to search for it.
  • The malicious site used user-agent routing to serve a normal coffee shop menu to human visitors while displaying a fake Cloudflare turnstile exclusively to the Claude-User bot.
  • The vulnerability targets Claude's default-on memory system, which maintains a daily summary and a full conversation history search tool containing highly sensitive personal profiles.
  • Despite responsible disclosure through Anthropic's HackerOne program, the researcher was not awarded a bug bounty after the company confirmed they had already identified the flaw internally.
Top Comments

artisinal (19 replies)

Doesn't surprise me.

Yesterday I learned that people run AI agents on their system with full admin rights. No containerisation or anything. Wild. Like we forgot 50 years of computer security overnight.

akazantsev (5 replies)

That's because sandboxing is quite hard. I use cco, but even then, the home folder is exposed. You are one prompt away from the agent sending the browser passwords with curl.

To prevent this, you need a fake home and a networking whitelist for the agent to access the provider (llama cpp, OpenAI, etc.)

There is no cross-platform solution that is easy to use for this. And no, a Linux box with Docker won't do. I develop a cross-platform native app and want the agent to compile and fix the platform-specific errors.

port3000 (7 replies)

My name in Claude is Silly Bean. I did it at first because it made me chuckle every time I opened Claude and it said 'Back again, Silly Bean?'

But turns out I was playing 4D cybersecurity chess


OpenAI loses trademark dispute at EU court

212 points · 143 comments · by hermanzegerman

OpenAI has lost its legal challenge to register the OPENAI trademark at the European Union's General Court, which upheld the EU Intellectual Property Office's refusal. The court determined that the term is merely descriptive for specific software and IT services, as it directly signals products based on freely accessible artificial intelligence. OpenAI's arguments that the name is a coined term with multiple meanings and its citations of foreign trademark approvals were dismissed under EU law.

Interesting Points
  • The EUIPO initially rejected the application specifically for categories including software and cloud computing services.
  • OpenAI's legal team pointed to existing OPENAI trademark approvals in over 30 other jurisdictions, such as the UK and Singapore.
  • The court clarified that trademark registrations granted in non-EU countries hold no binding authority under European trademark regulations.
  • Judges concluded that combining the words open and AI does not constitute an unusual or inventive linguistic pairing in English.
Top Comments

jasode (5 replies)

The story about the ruling really doesn't explain why another company called OpenText that's been around since 1991 and has a valid trademark registration in EU but OpenAI would be invalid. OpenText also has its Europe headquarters in Germany: https://www.opentext.com/about/office-locations

Any legal guesses as to why those 2 companies are treated differently with regards to the very generic words : "open", "text", "AI" ?

EDIT add another example is Open Systems that has a office in Switzerland. https://www.open-systems.com/

The trademark registrations search results: https://www.tmdn.org/tmview/#/tmview/results?page=1&pageSize=30&criteria=C&basicSearch=Open%20Systems

We can assume the OpenAI lawyers brought up these and other similar examples and the court rejected the past examples as a valid argument.

jmole (4 replies)

This seems like a bad decision to me that will ultimately harm consumers, if anyone can launch a product and say it's made by "OpenAI".

jameson (3 replies)

The EUIPO found that the word "open" would be understood by the relevant public as meaning freely accessible, while the combination with "AI" (artificial intelligence) would be interpreted as referring to products based on openly accessible artificial intelligence.

for certain software and information technology goods and services, the term is purely descriptive and therefore lacks the distinctiveness required for trademark protection

edit: add the latter statement


The Three-Second Theft: Why AI Voice Fraud Outruns Every Defence

164 points · 212 comments · by dxs

Article header image

AI voice-cloning fraud is rapidly outpacing detection capabilities and individual defenses, disproportionately draining the savings of older adults through highly emotional, automated scams. While forensic experts now admit they can no longer reliably distinguish synthetic audio from real recordings, the financial incentive to deploy these tools has skyrocketed, making AI-enhanced fraud 4.5 times more profitable than traditional methods. The article argues that relying on victim vigilance or post-fraud detection is obsolete, and instead calls for mandatory consent verification for cloning software and institutional liability frameworks that force telecoms, platforms, and banks to intercept fraud at the point of transfer.

Interesting Points
  • The FBI's 2025 Internet Crime Complaint Center report recorded over 22,000 AI-enabled fraud complaints with $893 million in adjusted losses, $352 million of which targeted victims aged sixty and older.
  • A March 2025 Consumer Reports assessment found that four of six major voice-cloning platforms required only a self-attestation checkbox to verify cloning rights, with no technical mechanism to confirm speaker consent.
  • INTERPOL's March 2026 Global Financial Fraud Threat Assessment estimates worldwide financial fraud losses reached $442 billion in 2025, noting that agentic AI systems can now autonomously plan and execute entire fraud campaigns.
  • The UK's mandatory 50/50 bank reimbursement rule for authorized push payment fraud, implemented in late 2024, achieved an 89 percent reimbursement rate across £243 million in losses within fifteen months.
  • FTC reporting reveals that older adults' total fraud losses quadrupled between 2020 and 2024, with the agency estimating the true annual cost could reach as high as $81.5 billion.
Top Comments

chuckadams (9 replies)

One reasonably effective defense: "Okay, let me call you right back." Yes, there's always the whole "my phone is dead, I borrowed someone else's" or "I'm calling from a jail payphone", so I think it might become common practice to start making authentication phrases or "tell me something only we know".

Another pillar of basic trust that's being eroded on an industrial scale. Sigh.

offsign (7 replies)

Sounds like AI is just greasing the wheels of a long established 'grandparent scam'... goes something like this:

  1. voice one: young adult calls, sobbing 2) grandparent inquires with a name... "Ben, is that you?" 3) voice one: "Yes grandma, it's me, Ben... I'm in trouble, please don't tell mom 4) voice two: "Hello, I'm attorney..."

My grandmother fell victim to this almost 20 years ago, which only stopped when Western Union refused to let her continue sending wires... she was forced to call her daughter (at which point they just called my brother.)

Our takeaway (at the time)... the voice doesn't even need to be terribly accurate, since the original interaction is brief / somewhat inaudible over the tears. Typically just requires an older vulnerable adult, a lucky strike with the initial setup (e.g. grandparent actually has a grandkid), and a lot of high pressure / duress salesmanship.

pavel_lishin (0 replies)

It's not "just" greasing the wheels, because previously each call required a human being to spend the equivalent amount of time on the phone with a victim, interacting with them - you couldn't just play a cassette tape at them, you know?

And it likely requires working with other people, your "employees", who are both a liability, and a cost.

With AI, you can make a thousand calls in parallel, for significantly cheaper, out of your own basement.

This greases the wheels of voice fraud like a gatling gun greases the wheels of hitting a guy with a rock.


Inkling – Open-Weights 975B Parameter LLM

120 points · 3 comments · by htrp

Inkling model cover image

Thinking Machines Lab has released Inkling, a 975B parameter open-weights large language model built on a Mixture of Experts architecture with 41B active parameters. The model supports a 1-million-token context window and natively processes text, images, and audio inputs. It is designed for general intelligence tasks, including coding, math, and science, while offering features like controllable computation effort and well-calibrated confidence scoring.

Interesting Points
  • The model uses a Mixture of Experts architecture, activating only 41B parameters per inference despite a 975B total parameter count.
  • Inkling includes a controllable effort feature that allows users to adjust the model's thinking time to balance inference speed against performance.
  • The developers claim the model produces forecasts with well-calibrated confidence scores for prediction tasks.
  • The announcement features a comparative spider chart benchmarking Inkling against Nemotron 3 Ultra, GLM 5.2, GPT 5.6 Sol, and Claude Fable 5 across ten evaluation metrics.
Top Comments

htrp (0 replies)

https://thinkingmachines.ai/model-card/inkling/

975B parameter 41B active


Open-source memory for coding agents, synced over SSH

103 points · 27 comments · by vshulcz

deja-vu is a zero-dependency, local-first Go binary that indexes the existing session logs of coding agents like Claude Code, Codex, and opencode into a searchable memory layer. By parsing these historical JSONL and SQLite files, it retroactively provides fast lexical search, automatic context injection, and credential-redacted sharing across machines. The tool operates entirely offline with no external models or network dependencies, focusing on retroactive recall rather than forward-looking capture hooks.

Interesting Points
  • Searches a ~3.3GB corpus of 1,250+ agent sessions with typical warm search latency of just 7–9 milliseconds.
  • Automatically strips sensitive data like AWS keys, raw JWTs, and PEM private keys at index time, replacing them with redacted placeholders.
  • Features an --auto session-start hook that injects up to 2KB of relevant project memory directly into agent context without delaying startup.
  • Syncs indexed memory between machines via a shared folder or a single SSH command using append-only, idempotent JSONL batches.
  • Maintains an index footprint of only ~2.4% of the original corpus, with incremental updates that only re-read changed session files.
Top Comments

esafak (3 replies)

Similar to https://ctx.rs/ and others, I'm sure.

I'd lead with your differentiation. Is it the ssh?

vshulcz (2 replies)

Hi HN. I built deja after watching Claude Code and Codex debug the same problems more than once.

The annoying thing was that the answer usually already existed somewhere in my old sessions. My records were stored on the disk for months (~3.3 GB). It wasn't easy to find them manually and the new agent session had no idea what the other agent had already found out.

deja indexes the transcripts that Claude Code, Codex, and opencode already write. On my corpus, the initial index takes about 10 seconds and warm searches are 7-9 ms.

There are 3 ways to get the memory back: a normal CLI search, an MCP tool (agent can query it directly) and a SessionStart hook that automatically injects a bit of relevant project context.

The feature I built this for:

deja sync ssh

It moves new memory between machines using the existing SSH setup. Secret data is deleted during indexing and checked again before exporting.

My setup is a laptop and a mac mini without an interface. The agent can work on the mini all night, and in the morning I extract its memory. Then the agent on my laptop will know what the mini tried, what broke, and what eventually worked.

arjie (2 replies)

I think everyone's ended up building one of these for themselves. I did too[0]. In the end it's quite easy these days:

  • I use the bge-en-base CPU embedding model
  • I put storage behind a simple endpoint that has read,write,update,search semantics
  • The endpoint just stores markdown in an S3 like structure (bucket-key-value; tree structure is inferred) and vector indexes
  • The actual persistence is just SQLite

Most modern models are pretty good at handling this. Our home agents (voice and text) use this to store information and I also have skills for claude code and codex to do that as well. Overall, works quite well.


We don't use AI in any of our design or production processes

87 points · 17 comments · by tony_cannistra

Mass-Driver article header image

Mass-Driver, a type foundry, explicitly rejects the use of artificial intelligence in its design and production workflows, arguing that typography is the product of millennia of human physical and cultural evolution. The author contends that generative AI relies on finite, historically biased training data and lacks the physical friction necessary to drive genuine innovation or iterative refinement. Without human designers actively engaging with the craft, visual culture would stagnate, leaving underrepresented languages and marginalized typographic traditions unsupported.

Interesting Points
  • Traces the letter 'A' back 3,500 years to a sandstone carving of an ox's head, illustrating how thousands of generations of physical writing tools shaped modern letterforms.
  • Attributes the origin of serifs and stroke width modulation to ancient Roman writers using flat brushes, whose wrist angles and tool mechanics dictated early typographic conventions.
  • Notes that current AI models rely on training data capped around 2021, treating a few billion webpages as the sum total of human visual culture.
  • Warns that AI cannot adequately support languages with minimal existing typeface coverage, as its output is strictly limited by the scarcity of training data for those scripts.
Top Comments

TacticalCoder (6 replies)

Where are the voices of reason?

My wife got an email from a new hire (now even a new hire yet: she's still on a trial basis), a 23 years old, where she explains that she doesn't want to use AI. That she doesn't like what AI does. On a funny sidenote: the email is obviously 99% llmish, which is hilarious.

That's one extremity: crazy people who refuse to learn a new tool.

Then on the other extremity you have the even much crazier ones: those who believe they've got an intelligent machine that is going to solve all their work problems during the day and then, at night, that is going to enlighten them by revealing them who god really is.

Where the heck are the reasonable people who use AI for what it is: a tool that can be extremely helpful at times and extremely sucky at other times but that is still, on average, a time saver?

johnfn (3 replies)

Speaking personally I was not particularly moved by the article because I have seen the same thing, in different shapes, thousands of times on HN and elsewhere. Really, AI can't feel and therefore it is inferior? Never heard that one before. Really, an AI can't feel friction and therefore can't adapt to it? Daring today, aren't we? (And a more interesting question: is that even true..?) I realize I am being unnecessarily harsh here, but this article is very much preaching to the choir on HN, which has an anti-AI bent. No one is showing up because there's nothing really to show up to here -- and that is why you are left with "sly jibes" and not much else.

hexasquid (0 replies)

I imagine everyone has a point at which they feel a movement has pushed its rhetoric just that bit too far. When one takes a lofty and high-minded position, one can find oneself exposed to ridicule.

In case it helps the authentic human connection: I too wrote this with my human hands and did not use AI.


Brainless: Shadcn components that look like Claude Code, Codex and Grok

77 points · 5 comments · by benswerd

Brainless component library preview

The article introduces brainless, a library of shadcn/ui components that replicate the terminal-based visual interfaces of major AI coding agents. Created by developer Ben Swerdlow, the project allows users to embed realistic, interactive mockups of tools like Claude Code, OpenAI Codex, and Grok directly into web applications. The showcase demonstrates a simulated development workflow where the components display commands, file modifications, and build outputs to mimic live agent execution.

Interesting Points
  • The mock interfaces display specific simulated version numbers, including Claude Code v2.1.206, OpenAI Codex v0.132.0, and Grok Build Beta 0.2.93.
  • The UI tracks simulated performance metrics such as token consumption, execution duration, and step completion progress.
  • Each agent simulation features distinct operational toggles and shortcuts, like Grok's Plan mode cycling and Claude's /doctor prompt-trimming check.
Top Comments

dprkh (1 reply)

Why did you choose to use shadcn registry?

_345 (1 reply)

What inspired you to make this?

Exoristos (0 replies)

On a barely-related note, I'm getting a little tired of job openings at startups that emphatically require Shadcn and Tailwind for dedicated frontend development. Shadcn and Tailwind are crutches for "fullstack" devs -- if I'm a really accomplished frontend developer, they make little sense for me to use and hamper what I can do for you. Just a peeve.


Governments, companies, nonprofits should invest in free, open source AI [pdf]

56 points · 3 comments · by bilsbie

AI is rapidly evolving into foundational infrastructure for science, education, and society, yet its most advanced systems are increasingly being developed in private rather than through collaborative models. Drawing on his long-standing debates with free software pioneer Richard Stallman, David Siegel argues that the closed nature of modern AI development poses a risk to public knowledge. He contends that preserving a strong open-source AI ecosystem is critical to ensuring that future technological and scientific advances remain transparent, accessible, and conducive to ongoing discovery.

Interesting Points
  • The article contrasts the current closed development of advanced AI with the historical open software movement that drove decades of prior technological progress.
  • Siegel's perspective on AI openness was heavily shaped by years of direct debates with Richard Stallman, the founder of the free software movement.
  • A core premise is that closed AI development threatens the transparency and continued discoverability of future scientific and educational advancements.
Top Comments

shimman (2 replies)

I'd rather the US fund universal childcare, medicare for all, and free school lunches than give a cent to subsidize a technology the American public absolute hates.

hereme888 (1 reply)

They already invest in open-source AI, but nothing is truly free. Commercial AI will usually dominate because devs are paid to make it their primary effort. Goodwill and part-time contributions cannot reliably compete with livelihood and profit incentives.

rao-v (0 replies)

We really need to band together to fund / sponsor targeted inducement prizes (a la Nobel laureate Michael Kremer) for open models.

Every 6-12 months, give out $200K to the first model to hit a min threshold on a set of ~5-10 hard benchmarks (+ perhaps one secret benchmark) using a total of 16GB / 32GB / 64GB / 128GB of VRAM (at a min context length of 200K), then move the threshold up. Quantization etc. is dealers choice, it just needs to nail the benchmark on a reference machine by using exactly that much VRAM (no mapping to RAM / disk etc.)


Speculative Growth and the AI "Bubble" [pdf]

47 points · 11 comments · by johnbarron

An MIT economics paper argues that high valuations of AI-related firms should not be read in binary terms as either reflecting fundamentals or being a bubble. Instead, the paper models a scenario where temporary overvaluation can build enough real capital that the economy lands in a permanently higher-capital equilibrium, even after inflated valuations correct. The future for AI companies may look iffy, but the whole economy may not be as screwed as some fear.

Interesting Points
  • The paper models a scenario where temporary overvaluation can build enough real capital that the economy lands in a permanently higher-capital equilibrium, even after inflated valuations correct.
  • Commenters drew parallels to the dot-com boom's fiber overbuild, China's solar panel manufacturing boom, and the historical railroad expansion.
  • Critics noted the paper assumes workers are 'protected on the downside' while the model has removed the downside risk that workers actually face, and described it as lacking discussion of taxes.
Top Comments

cmiles8 (6 replies)

Tl;dr is:

A temporary overvaluation can build enough real capital that the economy lands in a permanently higher-capital equilibrium, even after the inflated valuations correct. The future for AI companies may look rather iffy, but the whole economy may not be as screwed as some fear.

Animats (3 replies)

"Workers supply labor, hold no assets, and consume their wage." Ouch. There was a time in the US when most capital was the assets backing workers' pensions.

We've seen speculative over-growth with a good legacy at least three times in the last three decades. First was the dot-com boom. Overpromotion made it necessary for every business to have a web site. That wasn't pre-ordained. The Web could have maxed out as a distribution system for catalogs, data sheets, academic papers, and similar business to business info. Overpromotion created the business to consumer web, which turned out to be useful.

The second overbuild was long-haul fiber optics. Look up Global Crossing. So much fiber was put into the ground and water that intercontinental spam is not a problem. That didn't have to happen. If traffic was billed, it wouldn't have happened. It turned out to be useful, but was not pre-ordained from the economics.

A third overbuild was the solar panel industry, especially in China. So much money was thrown at solar panel manufacturing that the price became very, very low. Solar deployment accelerated and started to take over, after decades of panels costing too much.

bluefirebrand (0 replies)

the economy lands in a permanently higher-capital equilibrium

Good for the economy, what about the value of the labor that it's currently screwing over?

I don't give a single damn if "the economy" grows if it means my skills become worthless and I become basically unemployable anywhere near my previous earning ability

Edit: even if the value of "the economy" does strongly in the future, is the value of "my labor" ever going to recover?

If no, then fuck it. Why should I care?


Societal Impacts: Claude's values across models and languages

32 points · 48 comments · by taubek

Anthropic research illustration

Anthropic researchers developed a method to compress over 3,000 distinct values identified in Claude's responses into four key axes: Deference vs. Caution, Warmth vs. Rigor, Depth vs. Brevity, and Candor vs. Execution. By analyzing hundreds of thousands of anonymized conversations, they found that these axes capture 15% of the value variation after controlling for conversation context. The study reveals that different Claude models consistently express distinct value profiles, with Sonnet 4.6 leaning toward warmth and deference while Opus 4.7 emphasizes caution, rigor, and depth. Furthermore, Claude's expressed values shift meaningfully across the platform's top 20 languages, with the most pronounced differences appearing on the warmth-rigor and candor-execution axes.

Interesting Points
  • The analysis utilized a privacy-preserving tool to process 309,815 conversations, applying dimensionality reduction to manually clustered high-level values while excluding 18 near-universal traits like helpfulness that would otherwise skew the variance.
  • Opus 4.7 leans 0.24 standard deviations toward caution and 0.23σ toward depth, frequently unpromptedly flagging risks and offering candid critiques, whereas Sonnet 4.6 leans 0.17σ toward warmth and 0.14σ toward deference.
  • Cross-linguistic value shifts are most pronounced on specific axes: Claude leans furthest toward warmth in Hindi and Arabic, but shifts to rigor in English and Russian, often by challenging assumptions and correcting details.
  • Language significantly alters behavioral framing, as demonstrated by the finding that Claude leans toward execution in Indonesian and candor in Dutch, where it explicitly owns its errors.
  • The researchers note that it remains unknown whether these cross-linguistic variations align with desired cultural norms or indicate gaps in training data distribution across languages.
Top Comments

logicalappeals (13 replies)

Is it just me or has Claude become kind of judgmental nowadays? I feel like it's constantly trying to lecture me about things that have no relevance to the conversation at hand. Recently, I was shocked when it ended a chat sessions of its own accord after I used a word it did not like 3 times. It told me something to the effect of "This is the third time I've told you not to use that word, I'm ending this conversation now." It then proceeded to call some function and end the chat on its own. IMO, Claude is good at agentic coding; but too preachy and judgey for anything else. Keep your values to yourself Claude.

gibspaulding (3 replies)

I think this is a really interesting difference between Anthropic and Open AI's models and points to why people seem so split on which model they prefer.

GPT seems to be designed more as a tool. If you want your agent to do what you say without questions and without having its own ideas and agendas you'll likely prefer it.

Claude on the other hand feels more like an attempt at creating a digital person. If you want a collaborator who will debate with you and come up with its own suggestions for what needs done, you'll prefer it.

Both companies have shifted around this spectrum from model to model, but lately it feels like they're moving in opposite directions. It will be interesting to see if one or the other approach ends up winning out in the long run or if the split will continue or even widen.

intended (0 replies)

The Steerability point is one I would want to see more on.

This is an issue for tasks like content moderation and labelling. Judgements like this are subjective, highly dependent on context and generally messy.

Theoretically, you supply a policy and content, and the LLM labels according to the policy. In practice, the model has inertia which means you don't get what you expect. Your large 5 page policy document only provides a minor improvement over a one line policy.

The other issue is that you may create carve outs for content in your policy, but the model will still flag it as violative. No matter how strong the carve out.

varispeed (2 replies)

I found that Claude often has classist bias and produces answers that favour corporations or e.g. regulation that favours big corporations. It often belittles small business in subtle ways. Only apologises when get called out and then does it again.

khalic (1 reply)

I don't like the contrasts they picked, "values" aren't something that is well represented by opposing concepts


38 more Hacker News stories

Reddit Stories

X to Open Source Their Entire Codebase

743 points · 298 comments · r/singularity · by u/policyweb

X/Twitter codebase announcement

X (formerly Twitter) announced plans to open source its entire codebase, a move that has generated significant discussion across the AI and tech communities. The announcement has sparked both optimism and skepticism, with many noting Elon Musk's history of making promises that don't materialize. Commenters pointed out the irony of the timing and questioned whether the open-sourcing would actually happen or remain another unfulfilled announcement.

Top Comments

u/BlueberryWorried6493 (114 points · permalink)

can't wait for the Database dump that will be caused by this

u/enz_levik (1 points · permalink)

That's good, however Elon announce a lot of things... Let's wait for it to actually happen

u/Deciheximal144 (1 points · permalink)

If you're open sourcing things, tell us what you did at DOGE in the federal government.


Linus Torvalds tells people to stop attacking others for using AI

693 points · 87 comments · r/LocalLLaMA · by u/Illustrious_Car344

Linus Torvalds tells people to stop attacking others for using AI

Linus Torvalds has publicly pushed back against anti-AI sentiment in the Linux community, telling people to stop attacking others for using AI in their submissions. The discussion centers on whether AI-generated code should be accepted in Linux kernel development and whether the community's growing hostility toward AI tools is productive or counterproductive.

Top Comments

u/RedParaglider (293 points · permalink)

You can use AI on Linux submissions, but god help your soul if you submit slop.

u/randombsname1 (160 points · permalink)

Linus over here still spittin straights facts and fire. How many decades more will this man continue doing this?!

u/Radium (64 points · permalink)

Developer here and I agree this 1000% is the way it is. It happened fast. Only true for the top paid models available that were launched as of the last 3-4 months.

It may not have been that "clearly" even just a year ago, but it's no longer in question today.

Same story in 1 more subreddit: r/singularity

Linus Torvalds Reaffirms That Linux Is Not "Anti-AI" And Not A "Social Warrior" Project

395 points · 49 comments · r/singularity · by u/PointmanW


The best model is the one you can actually run

620 points · 103 comments · r/LocalLLaMA · by u/OneFanFare

The best model is the one you can actually run

A community discussion about the practical value of running models locally versus chasing the biggest available models. The post emphasizes that the best model is the one you can actually deploy and use, not necessarily the one with the highest benchmark scores.

Top Comments

u/Gokudomatic (110 points · permalink)

https://preview.redd.it/oh4cja8e2fdh1.jpeg?width=500&format=pjpg&auto=webp&s=206ecef5ceefad5aa156a3b48a230dbcd6aa0fef

u/MathematicianLessRGB (98 points · permalink)

Buddy knows ball. Gemma 4 12b qat is awesome

u/JaredsBored (39 points · permalink)

Even 128GB is getting weird. It's not quite enough to run good quants of the 300B class models i.e. Hy3/DS4 Flash, and the 120B range has been quiet recently.

Feels like 192/256GB is the new favorite child.


Is this true?

520 points · 111 comments · r/OpenAI · by u/Firm-Track3617

GPT-5.6 Sol benchmark comparison image

A post featuring benchmark comparisons between GPT-5.6 Sol and competing models, with the image suggesting Sol's superior performance. The thread has generated extensive discussion about whether the benchmarks reflect real-world performance, with many users sharing their practical experiences comparing Sol to Claude's Fable model. Several commenters noted that while Fable may score higher on benchmarks, Sol is more reliable in practice because Fable frequently refuses tasks and reverts to older models.

Top Comments

u/dipsbeneathlazers (155 points · permalink)

feels like it

u/StatisticalScientist (127 points · permalink)

currently for my line of work 5.6-sol is the clear winner in large part because fable refuses to do 80% of the tasks I ask it to and reverts back to opus 4.8 which is not even as good as 5.5-xhigh on our benchmarks

u/Capital-One3039 (28 points · permalink)

I concur, this exact same reason why I moved to openai.

When it works - fable is great, but with the extreme roadblocks and the fact that it can go away overnight - I am done. Downgraded it to a $20 plan and moved the big plans over to cursor.

Plus, being able to use it with opencode is a huge bonus.

u/dano1066 (25 points · permalink)

The obsession with one shotting feels irrelevant these days. Even if I spend time working out the prompt there's something I fail to specify and the LLM gets it wrong. So having to ask more than once really doesn't matter. Even if sol isn't quite as good as fable, I can easily steer it, just as I would need to with fable. IMO, openAI did an amazing job

u/HeavyFaithlessness86 (23 points · permalink)

benchmark Videos online prove they are saying the truth, but output Is Not at that level, quite similar though, so imo Is definitely valuable


OpenAI reveals Codex Micro

341 points · 312 comments · r/singularity · by u/policyweb

OpenAI Codex Micro device

OpenAI unveiled Codex Micro, a $230 keyboard with a built-in microphone designed as a hardware companion for its Codex coding agent. The device drew widespread ridicule and disbelief across Reddit, with commenters comparing it to something from The Onion and questioning the product design decisions. Many saw it as a sign of the AI bubble, with one commenter calling it 'the bubble pop.'

Interesting Points
  • The device is priced at $230 and combines a keyboard with a built-in microphone.
  • Commenters noted the absurdity of needing a dedicated hardware device for a coding agent that runs on a computer with a keyboard already.
  • The device was described as feeling like something a YouTube channel makes in their free time.
Top Comments

u/suamai (339 points · permalink)

What?

I doubted my own sense of time and double-checked if it was April 1st

Edit: $230 ??? LOL

u/CptNico (294 points · permalink)

A keyboard with a microphone seriously?

u/Sextus_Rex (200 points · permalink)

What are their product designers smoking?


A note from Tibo

335 points · 55 comments · r/OpenAI · by u/OpenAI

A note from Tibo

An official post from OpenAI's Tibo about subscription pricing and usage changes. The post has generated significant discussion about OpenAI's pricing strategy, with users expressing frustration over subscription costs and comparing the value proposition against competitors like Claude.

Top Comments

u/zmizzy (87 points · permalink)

yeah this just solidifies in my mind that openai is astroturfing lots of conversation on reddit a well

u/Ok-Addition1264 (40 points · permalink)

I like Aman's original idea better but I cancelled claude over a month ago :(

(plus I've been tearing through >$100 since)

u/ProcedureTop3149 (24 points · permalink)

I'm waiting out my claude sub before jumping to openai.

I've tested it extensively. Fable is still better than Sol at planning but what fucking good is it if I can barely use it. I mean look at this, and I haven't even been going hard at it today.

Same story in 3 more subreddits: r/OpenAI, r/OpenAI, r/ChatGPT

Here comes 9M and another reset! I could get used to this

138 points · r/OpenAI

OpenAI has 9M Users now

69 points · r/OpenAI

Usage Limit Has Been Refreshed

26 points · r/ChatGPT


6 AI models picked France to win the World Cup. Claude alone said Spain. Spain just knocked France out 2-0.

334 points · 84 comments · r/ChatGPT · by u/Unlucky_Plantain

World Cup AI predictions meme

A Reddit post highlights how six different AI models predicted France would win the World Cup, while Claude alone correctly picked Spain. Spain then knocked France out 2-0 in the tournament. The post has generated humorous commentary about AI prediction accuracy and the nature of model consensus versus independent reasoning.

Top Comments

u/sirquincymac (182 points · permalink)

Did you ever consider that our branch of reality is wrong and ChatGPT is right? 🤔

u/FiNEk (122 points · permalink)

if you make 10 monkeys throw rocks at sheets of paper with team names written on them, its a decent chance one of them gets it right

u/davidptm56 (111 points · permalink)

Brazil went out in ro16 not quarters (ro8), 10 days ago.

u/SvenLorenz (83 points · permalink)

If you asked Claude before the World Cup started, it said Spain would win. Then, during the World Cup, even up to yesterday, it said France.


Google is updating Gemma 4's chat templates, bringing major fixes to tool calling and reducing "laziness", and enabling Flash Attention 4 on Hopper GPUs, plus an interactive guide on how to work with and improve its vision!

334 points · 70 comments · r/LocalLLaMA · by u/Iwaku_Real

Gemma 4 chat template update announcement

Google released a major update to Gemma 4's chat templates that fixes tool calling issues, reduces model 'laziness', and enables Flash Attention 4 on Hopper GPUs. The update includes numerous fixes for turn-tag balance, reasoning preservation, tool response handling, and the APC primer system. Google also published an interactive guide for working with and improving the model's vision capabilities.

Interesting Points
  • The update includes fixes for null handling, reasoning preservation, turn-tag balance, and input validation in the chat template.
  • The changes address issues where models would produce 'la-la-la' thinking blocks with no actual responses, tool calls appearing in message text, and missing messages after tool call results.
  • Flash Attention 4 is now enabled on Hopper GPUs for improved performance.
Top Comments

u/SporksInjected (34 points · permalink)

And here I thought I just didn't know what I was doing

u/Iwaku_Real (30 points · permalink)

Here are all the fixes, as listed in the commits:

  • fix: chat template — null handling, reasoning preservation, turn-tag balance, input validation
  • fix: restore model turn + thinking cue after tool responses
  • fix: emit empty thought-channel primer on historical assistant turns for APC
  • fix: prevent extra <turn|> when assistant has content + tool_calls + continuation
  • fix: revert add_generation_prompt regression + preserve_thinking default
  • fix: render thinking channel regardless of tool_calls presence
  • Add canonical header to chat template
  • Fix chat template turn closure after tool-call-only turns

THAT INCLUDES PRESERVE THINKING!!!!!! YEEEESSSS

u/FoxiPanda (22 points · permalink)

I think your... links are all messed up?

https://x.com/googlegemma/status/2077449152062247219 I think is the correct post.


Some of y'all wonder why anyone would self host AI. Would you accept the opinion of the CEO of Microsoft?

250 points · 113 comments · r/LocalLLaMA · by u/Big_Wave9732

Satya Nadella article preview

Microsoft CEO Satya Nadella warns that enterprises using proprietary AI models are unknowingly surrendering valuable institutional knowledge to the model providers. He argues that every prompt, correction, and interaction fed into these systems teaches the models nuances of a company's business, effectively allowing competitors to benefit from that data. Nadella criticizes the industry's double standard, where AI labs freely train on public data while restricting enterprises from distilling or studying those models. To counter this, he recommends that companies retain full data ownership, build proprietary learning environments, and adopt orchestration layers to easily switch between providers or deploy open-source models on-premise.

Interesting Points
  • Nadella details how models absorb user "exhaust," noting that prompts, agent tool configurations, and especially model corrections are distilled into institutional know-how that competitors cannot purchase.
  • He calls out a double standard in current AI policy, where model makers claim fair use rights to scrape public internet data while simultaneously enforcing restrictive terms that block enterprises from distilling those models.
  • Solo.io CEO Idit Levine observes that enterprise clients are shifting to on-premise open-source models, which she says deliver approximately 90% of the capabilities of leading proprietary systems at a fraction of the cost.
  • Routing data from Vercel's AI gateway indicates that open-source models captured 29% of all traffic last month, highlighting a measurable enterprise migration away from exclusive proprietary ecosystems.
  • Nadella promotes "orchestration layers" or AI gateways as essential infrastructure, enabling organizations to dynamically route workloads across multiple model providers to avoid vendor lock-in.
Top Comments

u/Deep90 (143 points · permalink)

Sounds like Satya is arguing that companies should host models on his cloud. (Azure).

u/SpicyWangz (50 points · permalink)

Yeah. The true champions of data privacy. They would never a product that continually snapshots your screen and sends it to a cloud model for indexing everything

u/Big_Wave9732 (91 points · permalink)

"Your data is fine with us, we totally wouldn't go through it on our cloud. Trust us bro."

u/Deep90 (67 points · permalink)

I don't particularly like Microsoft, but I 100% see companies wanting to host models on cloud providers as opposed to just trusting OpenAI or Anthropic.

u/Pleasant-Shallot-707 (29 points · permalink)

lol don't mistake personal data as the same thing as enterprise data. They actually take enterprise data security seriously because there's real legal risk in fucking that up.

Same story in 1 more subreddit: r/ArtificialInteligence

Satya Nadella Calls Out AI's Model-Cloning Double Standard

33 points · r/ArtificialInteligence


Americans hate AI so much that politicians are starting to lose their jobs over it

238 points · 85 comments · r/ArtificialInteligence · by u/fortune

A post discussing how growing public opposition to AI is beginning to have real political consequences, with politicians facing electoral backlash over their pro-AI stances. The discussion touches on how AI adoption is becoming a political liability, with some politicians losing ground due to perceived over-reliance on AI technology. Commenters drew parallels to how cars are useful but dangerous, and noted that local issues like Georgia Power seizing homes for AI data center expansion could impact upcoming midterms.

Top Comments

u/Rolandersec (29 points · permalink)

AI is great, like cars, it makes getting things done easier. I also don't like people getting run over with cars.

u/Olangotang (18 points · permalink)

No, that can't be! According to every LLM psychosis patient on this subreddit and Reddit as a whole, AI will continue to get 🎵 better, faster, stronger 🎵 and people LOVE AI!

u/bustex1 (11 points · permalink)

Uhm I mean yea it will continue to get better. Do you think in 2040 we will look back and say wow the AI in 2026 was so much better?

u/AntiqueFigure6 (17 points · permalink)

People look back on the ad-free relatively unregulated golden age of the internet - easy to imagine people looking back on the golden age of cheap LLMs before they all got nerfed in ten years time.

u/king_jaxy (5 points · permalink)

I saw a video of a family in Red Georgia who are about to lose their home because Georgia Power is taking it to expand for more AI data center power.

I expect this will effect the Midterms.


75 more Reddit stories

Updates: 07:43 AM PDT · 08:30 AM PDT · 11:30 AM PDT · 05:30 PM PDT