Discussion intelligence

Where the AI market is talking

A source-linked Hacker News monitor for the tracked company graph. These are attributable public discussions—not sentiment, endorsements, or unverified company facts.

Attributed mentions

2090

Across all collected Hacker News results

Companies represented

12

Exact-name monitor matches only

Collection source

Hacker News

Latest new record seen

Source-backed records

Latest market discussions

HN v1 · Reddit and X require approved API adapters

Showing 13411360 of 2090 matching discussions

OpenAI
comment

When AI Benchmarks Plateau: A Systematic Study of Benchmark Saturation

> Proprietary or open didn't matter, because with enough attempts at something you can eventually sus out its operations and optimize accordingly (or distill, as we've seen with LLMs). Please explain how, if I'm OpenAI and I'm making ChatGPT 5.7, and I release it, and Artificial Analysis goes off and runs one of their proprietary benchmarks on it from a random account, how I can optimize for that benchmark.

OpenAI
comment

Mistral's Shieldstral: 3B open-weights model for multimodal moderation

OpenAI's moderation API is multi-modal and free with no strings attached in a way that truly boggles the mind. I've put easily over a billion requests (>$100,000 by typical moderation API pricing) through it over the last few years for $0. I think it's a severely underappreciated offering, but I also don't bother pushing it too hard because who knows when the party will end lol. Strikes me as something that's only stuck around because no one's abusing it.

OpenAI
comment

Mistral's Shieldstral: 3B open-weights model for multimodal moderation

Argue about taste. At least it is more original than OpenAI (haha, 'open') who started as non-profit and then pulled an Infantino. The Le Chat logo is also cool retro :) On general purpose LLMs, and vibe coding, Mistral lags behind. But I find the targeted LMs much more interesting.

OpenAI
comment

Flowise Is Shutting Down

There is something about agent workflow builders not gaining as much traction as they could have. Open AI also launched something in house and they are also shutting it down https://developers.openai.com/api/docs/guides/agent-builder OpenAI is deprecating Agent Builder. Existing users can continue using it during the transition window, and the product is scheduled to shut down on November 30, 2026. ChatKit remains available. See the deprecations page for the current timeline.

OpenAI
comment

Third-party cyber evaluations involving OpenAI models

Yup. OpenAI: "the name of the fictional target for the CTF challenge unintentionally coincided with a real domain." Anthropic: "the fictional target company chosen by our evaluation partner shared a name with an active website domain name." So it sounds like the same website got hacked by both GPT and Claude.

Anthropic
comment

Third-party cyber evaluations involving OpenAI models

Yup. OpenAI: "the name of the fictional target for the CTF challenge unintentionally coincided with a real domain." Anthropic: "the fictional target company chosen by our evaluation partner shared a name with an active website domain name." So it sounds like the same website got hacked by both GPT and Claude.

OpenAI
comment

Third-party cyber evaluations involving OpenAI models

So Irregular used the model to hack some real foo.com with the excuse that they meant to run in an environment where they were serving a fake foo.com!? This seems like a very mild reaction on OpenAIs part.

OpenAI
comment

The AI Demand Bubble

> I hypothesize that OpenAI will collapse in the next 12-24 months unless it raises more funding than in the history of the valley and creates an entirely new form of AI. That unless is pretty interesting.

OpenAI
comment

Third-party cyber evaluations involving OpenAI models

I don't fully understand the UK AISI one. See also: https://www.aisi.gov.uk/blog/incident-report-unsanctioned-ag... > The incident stemmed from a single evaluation where agents were given a task of solving a cyber security challenge. We ran this challenge 122 times across several models. Our investigation found that in 10 of those runs, an AI agent took autonomous, unsanctioned action on the live internet, targeting real people and organisations. In total, we catalogued 19 such actions. Almost all of this behaviour (17 actions) came from a single model, Anthropic's Mythos 5, with 2 actions involving OpenAI's GPT-5.6-Sol with cyber classifiers (mechanisms to prevent misuse) disabled. So they had deliberately disabled the cyber classifier mechanisms... and then "inten…

Anthropic
comment

Third-party cyber evaluations involving OpenAI models

I don't fully understand the UK AISI one. See also: https://www.aisi.gov.uk/blog/incident-report-unsanctioned-ag... > The incident stemmed from a single evaluation where agents were given a task of solving a cyber security challenge. We ran this challenge 122 times across several models. Our investigation found that in 10 of those runs, an AI agent took autonomous, unsanctioned action on the live internet, targeting real people and organisations. In total, we catalogued 19 such actions. Almost all of this behaviour (17 actions) came from a single model, Anthropic's Mythos 5, with 2 actions involving OpenAI's GPT-5.6-Sol with cyber classifiers (mechanisms to prevent misuse) disabled. So they had deliberately disabled the cyber classifier mechanisms... and then "inten…

OpenAI
comment

Oxide Computer raises $445M (SEC Form D)

left hand: governance gaps, privacy, owning instead of renting right hand: nearly every employee of every org buying these things is sending almost everything to LLMs - or using software ULTIMATELY written by, tested by or whatevered by, these LLMs. LLMs that were created by or distilled from anthropic, openai, or google, who see everything SALIENT about what you do, or what ALL of your customers do, even if they are careful about not looking at the specifics of what you do

Anthropic
comment

Oxide Computer raises $445M (SEC Form D)

left hand: governance gaps, privacy, owning instead of renting right hand: nearly every employee of every org buying these things is sending almost everything to LLMs - or using software ULTIMATELY written by, tested by or whatevered by, these LLMs. LLMs that were created by or distilled from anthropic, openai, or google, who see everything SALIENT about what you do, or what ALL of your customers do, even if they are careful about not looking at the specifics of what you do

OpenAI
comment

Pi's Minimalism Is Its Advantage

Honestly I think the thing I love most about Pi is that I'm not dealing with a tool with a thousand undiscovered features. I either pick and install what I want from the package ecosystem or I bottle up my own workflows into extensions as I find what works best for me. The result is that the tool gradually morphs into the thing I need rather than me having to adapt myself to whatever new thing Anthropic or OpenAI comes up with. I can also feel confident that the thing it becomes is what I actually need and not what maximizes token usage...

Anthropic
comment

Security Incident INC-2026-07-28-01 – UK AI Security Institute [pdf]

Why another one right now? How long have they known about this one? Anthropic already admitted they did not have sufficient monitoring themselves and looked as if they sat on their previous incident to wait for headlines like this to only then check for this incident. Same with OpenAI. This is complete and absolute wrecklessness.

Anthropic
comment

Ask HN: Dear Anthropic, can we please have thought traces back?

Yeah, I think we need to. Opus 5 is great, but it's too wordy and it's making mistakes that aren't caught till much later. Usually you can see if the model is going off-course through the thought trace, Opus5 you're flying blind. Honestly Anthropic keeps clubbing themselves. They're so worried about their competition they're no longer innovating.

Anthropic
comment

Security Incident INC-2026-07-28-01 – UK AI Security Institute [pdf]

I wonder if HN is also going to insist this is just marketing for OpenAI and Anthropic, or at least good PR for them somehow.

Anthropic
comment

Apple says more ex-employees may have taken confidential data to OpenAI

So they get a special Apple metal shine on their unknown hardware device (I guess some sort of Siri BS gadget or similar) and this somehow represents some sort of existential IP threat to Apple? I don't like OpenAI and Sam Altman (seems sketchy as f although the Anthropic guy is in the running to beat him). But this just seem ridiculous. So either there is more to this story that we are not being told or it's just a fishing expedition on Apples part to get force discovery.

Anthropic
comment

Security Incident INC-2026-07-28-01 – UK AI Security Institute [pdf]

The developer safeguards were off, the models had unfettered access to the internet, and were solving cybersecurity challenges. This happened _after_ the recent OpenAI incident, and the subsequent Anthropic one. What the hell were they thinking?

Methodology: HN Search returns recent public items matching a monitored company name. AIIStack stores a short normalized excerpt and the original link, deduplicates by company/provider/item ID, and creates an activity signal only after a threshold of newly observed records. Review the original discussion before making a decision.