Feed aggregator

Ask HN: Is there a test for measuring cognitive affects from LLM usage?

Hacker News - Thu, 09/03/2026 - 5:56pm

A part of me feels like I may be loosing my ability to think.

Is there a test I can periodically do to see how my cognitive abilities are being affected by my AI usage all day?

Comments URL: https://news.ycombinator.com/item?id=49557687

Points: 1

# Comments: 0

Categories: Hacker News

Nobody Is Saying Why OpenAI and Anthropic Had Outages Today

Wired Security - Thu, 09/03/2026 - 5:56pm
ChatGPT, Claude, and Grok all suffered outages at nearly the exact same time for reasons that remain murky.
Categories: Wired Security

Prediction Market Betting Is Getting People Banned and Arrested

Wired Security - Thu, 09/03/2026 - 5:48pm
This week on Uncanny Valley, we dig into the latest prediction market buzz, Flock’s AI-powered police search tool, and how tech bros don’t know how to talk about “rouge” AI agents
Categories: Wired Security

Both recent GCP outages caused by fiber optic maintenance

Hacker News - Thu, 09/03/2026 - 5:46pm

Both of GCP's recent outages were caused by "fiber optic maintenance". On August 20th when us-west1 went down they said [1]:

> The disruption originated during scheduled fiber optic maintenance, which unexpectedly compromised network capacity between data centers within the us-west1 region.

Then on September 1st when us-central1-b (and collateral affects in other zones) went down they said [2]:

> The immediate technical trigger for this event was the inadvertent physical disconnection of network fiber-optic cables during a routine hardware maintenance procedure.

Not sure if they're related but apparently procedures we're significantly changed after the first incident to prevent the second incident.

[1] https://status.cloud.google.com/incidents/utF3FMFdQfwBzJcGG6vf [2] https://status.cloud.google.com/incidents/J5ia5t9p3g9Q5Wi7r8Ev

Comments URL: https://news.ycombinator.com/item?id=49557563

Points: 2

# Comments: 0

Categories: Hacker News

WebMCP DJ – BananaLabs

Hacker News - Thu, 09/03/2026 - 5:41pm
Categories: Hacker News

CISO as a service, or CISOaaS, is the outsourcing of CISO (chief information security officer) and information security leadership responsibilities to a third-party provider.

Security Wire Weekly - Thu, 09/03/2026 - 5:05pm
CISO as a service, or CISOaaS, is the outsourcing of CISO (chief information security officer) and information security leadership responsibilities to a third-party provider.
Categories: Security Wire Weekly

Ask HN: Why don't LLM APIs have a first-class test mode?

Hacker News - Thu, 09/03/2026 - 4:59pm

Context: At work, we’re getting ready to stress-test a chatbot for scalability.

One fairly obvious issue came up: if our load tests exercise the real OpenAI/Claude APIs, a scalability test can quickly turn into a token-spending test.

Fair enough. We shouldn’t burn real inference just to test whether our own gateways, queues, WebSockets, streaming paths, retries, etc. can handle load.

The proposed solution was to mock all communication between our backend and the LLM provider.

Also reasonable.

What surprised me was the next step: we have to build and maintain that mocking service ourselves.

We can certainly do that. But should every company integrating with LLM APIs have to reinvent this?

Stripe solved a similar developer-experience problem years ago. They provide test mode, test data, test helpers, and even stripe-mock. It isn’t intended to perfectly reproduce Stripe’s backend behavior, but that’s okay. For many tests, you just need something API-compatible and predictable.

I’d love to see OpenAI, Anthropic, and other LLM providers offer something similar: an official API-compatible test endpoint that doesn’t invoke a model or consume billable tokens.

Ideally it could support things like:

* deterministic canned responses * streaming responses * configurable latency / time-to-first-token * configurable token counts * tool-call responses * 429s, 5xx errors and timeouts * malformed/interrupted streams * rate-limit simulation

The goal wouldn’t be to benchmark the LLM provider. You’d still need the real API for that. The goal would be to stress-test everything around the model without paying for thousands or millions of unnecessary inference calls.

What’s slightly ironic is that both OpenAI and Anthropic appear to use OpenAPI-based mock servers in their own SDK test suites. But, as far as I can tell, neither exposes that concept as a first-class public service for customers.

Am I missing something?

For teams running LLM applications at scale, how are you handling this today — homegrown mock server, generic HTTP mocking, record/replay, or just putting a budget cap on real API load tests?

(Comment is drafted and validated using ChatGPT Plus)

Comments URL: https://news.ycombinator.com/item?id=49556909

Points: 1

# Comments: 1

Categories: Hacker News

Pages