Attributed mentions
2195
Across all collected Hacker News results
Discussion intelligence
A source-linked Hacker News monitor for the tracked company graph. These are attributable public discussions—not sentiment, endorsements, or unverified company facts.
Attributed mentions
2195
Across all collected Hacker News results
Companies represented
12
Exact-name monitor matches only
Collection source
Hacker News
Latest new record seen
Source-backed records
Showing 361–380 of 2195 matching discussions
The knife is inanimate. The LLM is not. OpenAI prompted some employee to run the tests. The employee prompted the LLM. The LLM setup a message board and prompted other LLMs, and the fly wheel was running. It had to be turned off manually otherwise it'd still be going today. It's funny how a year ago talking about this kind of stuff would be laughed at by people like you, saying, "it's never happened before". Well it happened and you moved the goal posts like you always do.
How quickly this has changed since the time when people were boycotting OpenAI for making a deal with US Department of Defence and switching to Anthropic.
How quickly this has changed since the time when people were boycotting OpenAI for making a deal with US Department of Defence and switching to Anthropic.
Hint: It's not 15%. And there are many reasons, such as not having to keep up-to-date billing details in 70 providers, and not wasting money because most providers want you to prepay a balance that gets stuck in there if you switch to another provider. It also has way better uptime than the underlying platforms, even for proprietary models like Claude. When Claude APIs are having issues, OpenRouter Claude still keeps working because they can route to AWS Bedrock instead of Anthropic etc. This effect is even bigger with open-weight models because they typically have 5-10 providers.
> parroted in the media ugh just recently I saw a news article gushing about how AI had helped with some health/medicine study, and various patient organizations were all like "oh yeah this is a Good use of AI" and then I click the link to read the study and it's decision trees and clustering on a tiny dataset that you could analyze with a ten year old laptop. I mean, sure at some point decision trees and clustering were called "AI", but the way the article was written you'd think OpenAI and Anthropic were responsible for the advancement of medicine.
> parroted in the media ugh just recently I saw a news article gushing about how AI had helped with some health/medicine study, and various patient organizations were all like "oh yeah this is a Good use of AI" and then I click the link to read the study and it's decision trees and clustering on a tiny dataset that you could analyze with a ten year old laptop. I mean, sure at some point decision trees and clustering were called "AI", but the way the article was written you'd think OpenAI and Anthropic were responsible for the advancement of medicine.
Who is that someone? And unnamed employee that told somebody else? They make less than 140 million in revenue, what today is like 5 signing bonus for OpenAI :-) They are not profitable and see through less than 2 billion in revenue. These are easy to check so...
They might be popular for indie developers, but nobody doing serious AI or in a corporate environment is using them. And if they are, their compliance team is about to strike them down. The VCs forcing this acquisition do know this. - Why would you add a penalty of 50 ms at a minimum? And that is not the p95... Just run LiteLLM in house and you dont even really need that. - Their capacity pools are shared across the whole user base, a massive batch processing by another of their customers and think what that means for your response time... - So instead of negotiating corporate rates with OpenAI or Anthropic, you would be using an intermediary and topping up the corporate credit card... for a 5% markdown ? Really? - They can see all your critical corporate data on the in and out - They pre…
They might be popular for indie developers, but nobody doing serious AI or in a corporate environment is using them. And if they are, their compliance team is about to strike them down. The VCs forcing this acquisition do know this. - Why would you add a penalty of 50 ms at a minimum? And that is not the p95... Just run LiteLLM in house and you dont even really need that. - Their capacity pools are shared across the whole user base, a massive batch processing by another of their customers and think what that means for your response time... - So instead of negotiating corporate rates with OpenAI or Anthropic, you would be using an intermediary and topping up the corporate credit card... for a 5% markdown ? Really? - They can see all your critical corporate data on the in and out - They pre…
The other thing OpenRouter gets you is: Access to loads of models via a common API/sign-up/prepayment mechanism. I want to know which vendor/model does best at my extracting-facts-from-text task? Which does best at my OCR-a-text-document task? Which can deal with a safe-for-work beach photo without a censorship system false alarm? OpenRouter lets me run my tests against openai and anthropic and google and x and bytedance and qwen and llama, with a single sign-up and a single payment.
The other thing OpenRouter gets you is: Access to loads of models via a common API/sign-up/prepayment mechanism. I want to know which vendor/model does best at my extracting-facts-from-text task? Which does best at my OCR-a-text-document task? Which can deal with a safe-for-work beach photo without a censorship system false alarm? OpenRouter lets me run my tests against openai and anthropic and google and x and bytedance and qwen and llama, with a single sign-up and a single payment.
I used to demand the same but a word can have multiple meanings depending on the context, eg OpenMind means something specific. Open{$IT-term} does not necessarily mean Open Source {$IT-term} and OpenAI might mean an AI available to all people, not implying anything about cost. What could be illegal is the word Free, which a for-profit company would never use in the first place.
I very explicitly constructed the analogy to not require that! All a "cult leader" need be, in my analogy, is a passive question-answering oracle, the answers of which are biased by a semi-coherent preference function. The machine by itself is not an optimizer, certainly (much of that being by design—see various ~7-year-old conversations across the Internet about how to safely construct "tool AI", that has led almost directly to current model architectures.) But a bunch of mentally-ill people, who are indeed optimizers, can choose to allow the biases evident in the machine's output to become their own... and thereby effectively "bring to life" whatever partial echo of a will is recorded into the machine's output. Now, these same mentally-ill people could just-as-well do this with e.g. the…
Yes, but since we are specifically talking about claude.md, Anthropic themselves claim on their website these files "serve[s] multiple purposes: providing architectural context, ..." ( https://claude.com/blog/using-claude-md-files ) They also recommend starting with an /init command, which is also something the study very specifically called out as "having a marginal negative effect" And from personal experience, I can only confirm that many people seem to see this as the main purpose of agents/claude md files - a persistent architectural overview of your project. Quick Edit: My point simply being that I think it's understandable if some people don't understand the big deal about these files because they've had a drastically different experience than other pe…
I don't trust anything OpenAI says anymore.
Great to hear that you are planning larger benchmarks! I am particularly interested in longer-running tasks with many steps and self-correction. Divergence is fine as long as the model can still solve the task, which Divergence-300 @32 does not measure. The current benchmark suites that frontier AI labs use are probably a good fit, e.g. https://z.ai/blog/glm-5.3#:~:text=Performance%20across%20com... https://www.kimi.ai/ai-models/kimi-k3#:~:text=Performance%20... https://www.anthropic.com/news/claude-opus-5 https://openai.com/index/gpt-5-6/ But guessing from your current benchmarks, I assume that you are severely compute-constrained. What is your time budget?
> I'm not sure why these types of news are still coming out when we all have AI at work. Because OpenAI is burning $15 billion/year, and outrageous stories like those get parroted in the media. It's free marketing for a company desperate to get middle management to believe that a $500/mo subscription is absolutely crucial for every single employee.
> but very few of them are cyber focused, so it’s not surprising IMO Yeah but that's a business decision. I work on security at a company that does sandboxing and when the company decided to build an AI harness I was brought in as one of the earliest engineers on the product. We do almost everything on that list and we're a fraction of the size of OpenAI. And it wasn't particularly hard, and we have harder requirements imo (because we solve more general problems vs "run a very specific agent with a very specific task and very specific access"). > And you’re also not fully considering the granularity problem, eg there are a lot of sandboxing tools out there but they’re usually quite coarse in the dials and levers they offer, so the only way you can still make the workload do what it …
That's incorrect. My thought process hinges upon the fact that Hugging face did not report that uncontrolled models escaped confinement autonomously by coordinating with other models via a series of zero days during training. OP is concerned about the autonomous nature of the incident, not whether or not the incident happened. Nor am I contesting that the incident happened.
I see. Following your conjecture, there are two possibilities: 1. It wasn't an accident. OpenAI explicitly directed its agents to hack Hugging Face. Despite the fact that such a thing is a federal crime that carries prison sentence. 2. It wasn't an accident. OpenAI and HuggingFace conspired and let the hack happen for publicity. Is there anything I'm leaving out?
Methodology: HN Search returns recent public items matching a monitored company name. AIIStack stores a short normalized excerpt and the original link, deduplicates by company/provider/item ID, and creates an activity signal only after a threshold of newly observed records. Review the original discussion before making a decision.