High-Throughput Continuous Batching & Paged KV Cache Engine (1-vCPU Benchmark) - bmartin-systems/cortex-serving-arena-preview
emschwartz's feed
Why MCP was always a bad idea and I'm tired of pretending it's not.
28 tracks, 343+ hands-on challenges. Implement Raft, CRDTs, sharding, and storage engines in your browser in 8 languages.
'Jev' doesn't chat. It produces typed probabilistic decisions
Covered by17 sites
Medium, Stackness, and more ›
Tesla did not tell the city transportation department it had deployed Cybercabs in the five boroughs. [more ›]
A type foundry for bastard web fonts.
Okay, I did not actually forget, I am very much aware of their existence, but I realized something that is very obvious in hindsight
Covers 1 story
GCVE Workshop Before The Vulnopticon Conference
A public timeline of notable AI hacking and AI-enabled cyber incidents.
DQuic is an open-source QUIC implementation purpose-built for DHttp.
The September 2026 DuckDB Ecosystem Newsletter: DuckLabs joins AWS and MotherDuck acquires Tower, DuckDB v2.0-alpha lands, table functions in pure Java, bulk loads into SQL Server, Zarr stores as SQL tables, an infinite canvas for your data, plus a community spotlight on Vladimir Gribanov.
Covers 4 stories
This paper analyzes the amount of source code in GNU/Linux, using Red Hat Linux 7.1 as a representative GNU/Linux distribution.
Covers 1 story
Learn how to write a spec, turn it into a plan, and let Claude Code build using spec-driven development. Compare Superpowers, Spec Kit, and BMAD-METHOD.
Covers 1 story
Hi r/rss, Like many of you, I love RSS because it puts us back in control of our information diets. But lately, I found myself getting overwhelmed by news fatigue—wading through 20-paragraph articles filled with narrative padding, cookie banners, and paywalls just to find 2 or 3 core facts....
Get API key Read the docs See top models, labs, apps, and providers on AI Gateway with token, request, and cost trends. ## Copy link to headingToken Volume Share of AI Gateway token volume across the top models. Stacked chart Line chart Jun 21, 2026Sep 18, 2026 10 Other 16.5% ## Copy link to headingOpen vs. Closed Token Volume Share of AI Gateway token volume on models that publish their weights for download, versus everything else....
Build a GitHub Actions workflow that runs quality checks, caches work with Turborepo, and enforces Lighthouse performance budgets.
Two Windows recovery and deployment DLLs have shipped Rust since the first 24H2 build and are now 96 percent Rust by the linker's own accounting. A census of 13,988 live files and 44,889 more in the component store on one Windows 11 25H2 installation, with Microsoft's public symbol files used to measure how much of each binary rustc produced, and 83 GPU driver matches explained by LLVM rather than Rust. Hashes, symbol tables and the full census attached.
Covers 1 story
We begin with a bird’s eye view of some important aspects of intuitionistic type theory. Readers who are unfamiliar with the theory may prefer to skip it on a first reading.
RetinaFace is an advanced AI face detection model that delivers accurate facial landmark recognition, face alignment, and facial analysis for computer vision applications.