Feed aggregator
Ask HN: Is there a test for measuring cognitive affects from LLM usage?
A part of me feels like I may be loosing my ability to think.
Is there a test I can periodically do to see how my cognitive abilities are being affected by my AI usage all day?
Comments URL: https://news.ycombinator.com/item?id=49557687
Points: 1
# Comments: 0
Nobody Is Saying Why OpenAI and Anthropic Had Outages Today
CERN moves accelerator control computers to Debian
Mullvad migrating to Q9 for encrypted DNS
Article URL: https://mullvad.net/en/blog/shutting-down-our-public-encrypted-dns-servers-and-sponsoring-quad9-instead
Comments URL: https://news.ycombinator.com/item?id=49557666
Points: 2
# Comments: 1
Court Rules Against Citizen Journalists in DMCA Takedown Case–EFF Will Appeal
Article URL: https://www.eff.org/deeplinks/2026/09/court-rules-against-citizen-journalists-dmca-takedown-case-eff-will-appeal
Comments URL: https://news.ycombinator.com/item?id=49557653
Points: 2
# Comments: 0
I built an AI development environment that runs and builds software on Android
Article URL: https://github.com/sedds89/StitchLabtools
Comments URL: https://news.ycombinator.com/item?id=49557621
Points: 1
# Comments: 0
I used AI to connect with 30 people at a conference
Article URL: https://galvered.com/blog/ai4-conference/
Comments URL: https://news.ycombinator.com/item?id=49557613
Points: 1
# Comments: 0
Prediction Market Betting Is Getting People Banned and Arrested
Both recent GCP outages caused by fiber optic maintenance
Both of GCP's recent outages were caused by "fiber optic maintenance". On August 20th when us-west1 went down they said [1]:
> The disruption originated during scheduled fiber optic maintenance, which unexpectedly compromised network capacity between data centers within the us-west1 region.
Then on September 1st when us-central1-b (and collateral affects in other zones) went down they said [2]:
> The immediate technical trigger for this event was the inadvertent physical disconnection of network fiber-optic cables during a routine hardware maintenance procedure.
Not sure if they're related but apparently procedures we're significantly changed after the first incident to prevent the second incident.
[1] https://status.cloud.google.com/incidents/utF3FMFdQfwBzJcGG6vf [2] https://status.cloud.google.com/incidents/J5ia5t9p3g9Q5Wi7r8Ev
Comments URL: https://news.ycombinator.com/item?id=49557563
Points: 2
# Comments: 0
GitHub dipped 7%, while both AI labs were down
Article URL: https://claude.ai/public/artifacts/2368658e-60f2-4f51-8883-8d72738ba7ce
Comments URL: https://news.ycombinator.com/item?id=49557544
Points: 4
# Comments: 1
Watermarks Track AI Generated Content [video]
Article URL: https://www.youtube.com/watch?v=kVXp6UNVPTo
Comments URL: https://news.ycombinator.com/item?id=49557525
Points: 2
# Comments: 0
The $1 Trump Coin
Article URL: https://www.theguardian.com/us-news/2026/sep/02/trump-one-dollar-coin
Comments URL: https://news.ycombinator.com/item?id=49557519
Points: 2
# Comments: 0
GPT-6 Astra: an automated AI Engineer you can hire for <$6 an hour
Article URL: https://www.latent.space/p/astra
Comments URL: https://news.ycombinator.com/item?id=49557493
Points: 2
# Comments: 0
WebMCP DJ – BananaLabs
Article URL: https://devpost.com/software/bananalabs-webmcp-dj
Comments URL: https://news.ycombinator.com/item?id=49557486
Points: 1
# Comments: 1
Daybreak for Frontline Defenders: $1B to protect essential services
Article URL: https://openai.com/index/daybreak-for-frontline-defenders
Comments URL: https://news.ycombinator.com/item?id=49557449
Points: 1
# Comments: 0
CISO as a service, or CISOaaS, is the outsourcing of CISO (chief information security officer) and information security leadership responsibilities to a third-party provider.
Online Sports Betting Is a Public Policy Issue and a Public Health Issue
Article URL: https://jamanetwork.com/journals/jama-health-forum/fullarticle/2845354
Comments URL: https://news.ycombinator.com/item?id=49556924
Points: 1
# Comments: 0
Ask HN: Why don't LLM APIs have a first-class test mode?
Context: At work, we’re getting ready to stress-test a chatbot for scalability.
One fairly obvious issue came up: if our load tests exercise the real OpenAI/Claude APIs, a scalability test can quickly turn into a token-spending test.
Fair enough. We shouldn’t burn real inference just to test whether our own gateways, queues, WebSockets, streaming paths, retries, etc. can handle load.
The proposed solution was to mock all communication between our backend and the LLM provider.
Also reasonable.
What surprised me was the next step: we have to build and maintain that mocking service ourselves.
We can certainly do that. But should every company integrating with LLM APIs have to reinvent this?
Stripe solved a similar developer-experience problem years ago. They provide test mode, test data, test helpers, and even stripe-mock. It isn’t intended to perfectly reproduce Stripe’s backend behavior, but that’s okay. For many tests, you just need something API-compatible and predictable.
I’d love to see OpenAI, Anthropic, and other LLM providers offer something similar: an official API-compatible test endpoint that doesn’t invoke a model or consume billable tokens.
Ideally it could support things like:
* deterministic canned responses * streaming responses * configurable latency / time-to-first-token * configurable token counts * tool-call responses * 429s, 5xx errors and timeouts * malformed/interrupted streams * rate-limit simulation
The goal wouldn’t be to benchmark the LLM provider. You’d still need the real API for that. The goal would be to stress-test everything around the model without paying for thousands or millions of unnecessary inference calls.
What’s slightly ironic is that both OpenAI and Anthropic appear to use OpenAPI-based mock servers in their own SDK test suites. But, as far as I can tell, neither exposes that concept as a first-class public service for customers.
Am I missing something?
For teams running LLM applications at scale, how are you handling this today — homegrown mock server, generic HTTP mocking, record/replay, or just putting a budget cap on real API load tests?
(Comment is drafted and validated using ChatGPT Plus)
Comments URL: https://news.ycombinator.com/item?id=49556909
Points: 1
# Comments: 1
