evaluating

8 stories and discussions about evaluating, aggregated from every source we track.

1.

Evaluating Jev's calibration on two hard datasets.

3 points•cannedbread•4 days ago•0 comments•
2.

SLEEF implements vectorized C99 math functions.

2 points•fanf2•4 days ago•0 comments•
3.

I’ve spoken several times about jittered Voronoi grids, such as here. These are infinite Voronoi diagrams where the sites come from randomly picking a point from each square in a square grid.…

1 points•ibobev•3 days ago•0 comments•
4.
1 points•mooreds•5 days ago•0 comments•
5.

Testing smolvm 1.8.3 shows it is well suited for sandboxing untrusted Python and JavaScript data transformations using hardware-isolated VMs rather than shared-kernel containers. Offline local images, no-network…

1 points•binsquare•8 days ago•0 comments•
6.

From “Storing” to “Staying Current”: Why Agent Memory Needs a Shared Evaluation

1 points•IreneAI•9 days ago•0 comments•
7.

AI-generated music detectors are commonly compared using aggregate scores on benchmarks whose training overlap, generator lineage, source provenance, and audio-transformation history are only partially observable. This…

1 points•unohee•9 days ago•0 comments•
8.
1 points•IreneAI•10 days ago•0 comments•

Related topics