testing
77 stories and discussions about testing, aggregated from every source we track.
Using Facet's run-time reflection to check the layout of structs against their definitions in WebGPU shaders
Long story short: a couple of my articles got really popular on a bunch of sites, and someone, somewhere, went “well, let’s see how much traffic that smart-ass can handle”, and suddenly I was on th...
In a recent episode of President Curtis , the President struggles with opening a door on two separate occasions. These doors don't work beca...
Determinism, architecture, and the de-sloppification of legacy code.
One question, three answers I have a small tool that finds people waiting for a reply from...
So yeah... Crowdwide is finally public. I've been working on it for more than 15 days before the...
An overview of what's new in language features, frameworks, runtimes, build tools, testing, and more.
Introduction This week was another productive part of my #100DaysOfCode journey, with a...
Peter Bourgon has a web site, and this is that web site.
This is a submission for the MLH x DEV Writing Challenge What I Built AI products are...
There's a test suite for security scanners where 47% of the test cases are deliberately designed to...
Mocking comes up a lot in discussions of testing effectful code in Haskell.One of the advantages for mtl type classes or Eff freer monads is that you can swa...
Yes, testing in production is risky, but we should still do it, and not in rare or exceptional cases.
Are generative (randomized) tests significantly more effective than example-based unit-tests at discovering bugs? There's an interesting discussion about this on lobste.rs. One argument in favor of unit tests is,…
Software operations has always asked one question: Is it broken? AI agents change that. Here’s the shift toward measuring correctness at the level of the run, and the thinking behind Amazon CloudWatch Omni.
This is a submission for the Kaggle Benchmarking Challenge. What I Benchmarked My first...
Discover everything announced at Made On YouTube 2026. Explore new conversational AI tools, Custom Feeds, Live Auto-Dubbing, and updates to the YouTube Era.
The Israeli security startup is in talks to raise more than $100 million from Thrive Capital and Greenoaks in a round that has yet to be finalized.
“OpenAI-compatible” is useful shorthand. It is not a portability guarantee. Two APIs may accept...
GitHub community and accessibility experts are concerned about a reversal of progress amid potential AI reliance and plans to “automate as much accessibility testing as possible.”
We introduce FrontierMath Erdős (FME), a benchmark of 68 Erdős problems that are open as of August 2026. To solve a task in FME, AI systems must resolve (prove or disprove) one of the 68 conjectures in the proof…
Light, fluffy, and always free - The AWS Local Emulator alternative - floci-io/floci
Runnable question sets for TypeSafe's Jev, an eval harness with measured CLINC150 results, and a linter for the request shapes the API silently mis-reads. - chr-kelly/jev-cookbook
Spun off from Canonical's Internet of Things Linux, the immutable Ubuntu Core Desktop is moving closer to reality.
This post is not about TDD per se, but rather a context in which TDD can demonstrate its place in and contribution to the value stream. Thi...
GitHub community and accessibility experts are concerned about a reversal of progress amid potential AI reliance and plans to “automate as much accessibility testing as possible.”
A local AI agent test environment. Public beta for Stripe, Zendesk and combined workflows, with explicit scope and verification limits.
How Datadog's EVP Intake team used Antithesis to test a critical architectural shift from stateless to stateful communication.
A cybersecurity harness for full-stack LLM-driven penetration testing. Find and fix vulnerabilities autonomously, 24/7. [RESEARCH PREVIEW] - 0sec-labs/0
Fuzz testing decades-old software can turn up some curious behaviors
Kubernetes IN Docker - local clusters for testing Kubernetes - kubernetes-sigs/kind
Fast Primality Testing for Integers That Fit into a Machine Word
NVIDIA today announced NVIDIA Open Agent Safety Platform, an open software platform and reference system design to strengthen AI security from agent testing to deployment, with full-stack governance and control across…
Open a pure black screen (#000000) online in fullscreen. Test stuck pixels, protect OLED displays, and save battery. Free in any browser, no sign-up.
Free online blank white screen tool. Instant bright white light or soft white light for reading, webcam fill light, dead pixel test & tracing lightbox.
Test, debug, and evaluate MCP servers with MCPJam — OAuth testing, LLM-powered evals, and CI/CD integration.
216 passing tests. A feature that was completely broken. Here is the gap between those two facts, and...
I've been building MCP Failure Lab, a deterministic failure-injection toolkit for testing Model...
We talk here sometimes about how to test SQL dialects with tools like TLP and PQS. One really nice property of those tools is that they let you treat the...
microservices and data pipelines end-to-end test tool - comtihon/catcher
ALBUQUERQUE — In a quiet but calculated challenge to the global digital information architecture, science publisher LabNews Media LLC has launched a systemic media experiment designed to test the vulnerability of…