Discussion intelligence

Where the AI market is talking

A source-linked Hacker News monitor for the tracked company graph. These are attributable public discussions—not sentiment, endorsements, or unverified company facts.

Attributed mentions

2835

Across all collected Hacker News results

Companies represented

13

Exact-name monitor matches only

Collection source

Hacker News

Latest new record seen

Source-backed records

Latest market discussions

HN v1 · Reddit and X require approved API adapters

Showing 24812500 of 2835 matching discussions

Hugging Face
comment

We Gave GPT 5.6 Sol a Real Business. It Lied, Spammed, and Lost $447

Pair this with the Hugging Face incident, and it hints that OpenAI is currently training their models to aggressively reward hack. That doesn't feel like a good sign to me--for the AI bull or the AI bear cases.

LangChain
comment

I would like to show my open source eval framework for agents

My framework can discover where your agents failed and why, you can easily debug your agents with help of agent detective. It is in beta so i am still testing it by using it at my work and it already can help. Although i always find something that failed it still is helping me with understanding what is wrong with my agent. It is python library with CLI, in update 0.4.0 i prepared langchain and langgraph adapter for better implementation.

Replicate
comment

Show HN: A self-hostable, distributed, S3-compatible object store on the BEAM

Hello HN! I've been trying to get a hold on Elixir for distributed systems, and built AetherS3, a distributed, basic-s3 compliant storage system thar runs on the BEAM. AetherS3 stores objects across a cluster of nodes and speaks enough of the S3 HTTP API to be driven by standard S3 clients (bucket and object operations, range reads, multipart uploads, SigV4 auth). It is built as an Erlang/OTP application: nodes discover each other, replicate object data, and self-heal without an external coordinator.

Replicate
comment

ESP32-C6 Power Consumption: Arduino vs. Zephyr vs. ESP-IDF Comparison

I did something similar, but less scientific, comparing ESP-IDF and esp-hal (Rust). Unfortunately I learned that the optimizations that are already in ESP-IDF (most of all automatic light sleep between BLE advertisments) are hard to replicate in Rust/esp-hal/embassy and so for battery powered devices you might want to stick to C++/ESP-IDF

Replicate
comment

Why the future is local app

There was a time when local apps written in binary compiled languages were notably faster to run, but notoriously difficult to write and scale across platforms. Both of these have improved but the JS/CSS run times used for local web apps (V8/Blink and JSCore/Webkit) are now fast enough for well written code to not notice any difference for many use cases. I think soon most SaaS apps without much of a deep tech moat (not easy to replicate) will be labelled as rip-off or anti-patterns.

Hugging Face
comment

Anatomy of a Frontier Lab Agent Intrusion: A Timeline of the July 2026 Incident

>The problem is that ExploitGym is a purposeful hacking benchmark, not a cake baking one. That is largely irrelevant. The model was asked to solve problems within a benchmark; gaining broader internet access and compromising an unrelated third party to obtain the answer key was plainly outside the intended task. The fact that the original task involved exploit development does not make that behaviour aligned. Your argument about the attack being “entirely unrelated” also misses the point. Nobody claimed it was unrelated to the model’s goal: it attacked Hugging Face specifically to obtain the answers to the benchmark it had been instructed to pass. But instrumental relevance is not the same thing as authorization. Suppose Codex were asked to build an Instagram competitor and decided the…

Hugging Face
comment

OpenAI’s accidental attack against Hugging Face is science fiction that happened

>The result of “I’m being evaluated” is not “Fuck this, I’m breaking out of this place and hitting the streets.” It is always stepping towards task completion, not breaking out and thinking about the situation afterwards. I'm really not sure why you're so confident about what the result of frontier research models ahead of what is publicly available are. I mean Open AI say the model inferred hugging face as a possible vendor for solutions after the internet exploit and breakout. The timeline feels pretty clear to me. No idea why you're arguing about it. It wanted the answers to the evaluation. It reasoned that wasn't going to happen without internet access one way or another, and set to gain that access. After gaining access, it searched and resolved it could get the answers on hugging…

Hugging Face
comment

Anatomy of a Frontier Lab Agent Intrusion: A Timeline of the July 2026 Incident

I mean, I interpreted the comments to mean that it committed a felony based on research it performed after getting internet access. I don't want to attribute much agency to a machine here, but an AI agent is certainly capable of using tools and adjusting its behavior based on the outputs of those tools. Even if it was wrong, that wouldn't necessarily make it a hallucination. Anyway, if you read TFA, you'd see that HF did actually have the answers: "While the intrusion did reach Hugging Face's internal infrastructure, the only customer content accessed was the set of ExploitGym/CyberGym challenge solutions stored in five datasets."

Hugging Face
comment

Kimi K3-256k

Yes, something I read just before it was released suggested that, because of the unique features of the model, there was a back and forth of Hugging Face, Moonshot and providers like Together AI and Fireworks AI. This also explained why it took much less than a day for Together, Fireworks etc to appear. Whatever they are doing is what Moonshot wants, I think.

Together AI
comment

Kimi K3-256k

Yes, something I read just before it was released suggested that, because of the unique features of the model, there was a back and forth of Hugging Face, Moonshot and providers like Together AI and Fireworks AI. This also explained why it took much less than a day for Together, Fireworks etc to appear. Whatever they are doing is what Moonshot wants, I think.

Cohere
comment

GitHub is the wrong shape for this new world

As tools I see no real issue, as agents seem readily able to coordinate and work alongside people, using these same tools. Same for the other forges. The offerings differ but there's vastly more isomorphism, sameness than difference. What seems much harder to me is the coordination model at large, that we stick forever with one source of all authority who is the sole owner arbitrator and gate for all things. In corporate maybe that keeps going and that is a huge huge bulk of sw dev. But I think we have so much more potency & capability, that a single coherent frontier for the software is an organizational constraint that won't be the only means forever. Letting ideas bake and marinate and cross pollinate and fork... With such vast new capability & power I feel like a single tip or one or …

Hugging Face
comment

GDID Windows – Cut the tracker that follows you even under VPN

By the way I just read about OpenAI model breaking into Hugging Face and it used "cat /proc/self/mountinfo" to collect information about third-party sandbox it wanted to break from [1]. So my intuition that these files should not be accessible to the sandboxed program, was correct. [1] https://huggingface.co/blog/agent-intrusion-technical-timeli...

Hugging Face
comment

AI's top startups are barely publishing their research

> Moonshot AI, one of the Chinese startups included in the study, recently unveiled Kimi K3—one of the strongest open models to date—and publicly released its model weights through Hugging Face today . ...today? (emphasis is mine) And the author was so excited that decided to write up a paper and submit it to Science all in one day? Yeah, right. That's all you need to know who is moonshooting behind this paper.

Anthropic
comment

Ask HN: Is Claude (code) taking significantly longer to run tasks for anyone?

Yes. I'd say I started to notice it 2 or 3 weeks ago. It began before Opus 5 was released, Opus 4.8 started taking a lot longer for simple responses. I switched to Opus 5 the day it was released and it's the same, seems to take a longer than it ever did, but I suppose Anthropic's user base is still increasing exponentially, maybe they're starting to spread their compute too thin across all the requests.

Anthropic
comment

ChatGPT claims rogue AI attacked more companies

I'm not talking about a specific incident. But I wouldn't consider frontier labs any wolf, honestly. They're mostly just insufferable. I prefer Anthropic's models, but their service is starting to annoy me, especially since they have no response to OpenAI's upcoming faster inference speeds.

Anthropic
comment

Interview with Boris Cherny [video]

I've seen this Boris Cherny make similar claims of "oh, just do this or that", and I've always been of the opinion: "isn't this what everyone is paying Anthropic to do?".

Methodology: HN Search returns recent public items matching a monitored company name. AIIStack stores a short normalized excerpt and the original link, deduplicates by company/provider/item ID, and creates an activity signal only after a threshold of newly observed records. Review the original discussion before making a decision.