SWE-Bench Multimodal: Do AI Systems Generalize to Visual Software Domains?swebench.com·1 pts·matt_d·0
FLARE: Verifying MILP Reformulations with LLM-Based Theorem Provingflare.henryrobbins.com·2 pts·matt_d·0
CascadeLUT: Info.-Ordered Streaming Inference for Bandwidth-Constrained FPGAsarxiv.org·1 pts·matt_d·0
Trustworthy Human-AI Collab in a Live Type-Theoretic Computational Commonsjanestreet.com·2 pts·matt_d·0
Computing Against the Odds: Processor Architectures for Batteryless Devices [video]youtube.com·1 pts·matt_d·0
AutoResearchExam: Measuring agents' ability to improve and generalizebenchmarks.bespokelabs.ai·3 pts·matt_d·0
Performance Foundations of Parallel and Distributed Reasoning Language Modelsarxiv.org·1 pts·matt_d·0
Testing race conditions with memory access tracing & stack-based delay injectionprojectzero.google·4 pts·matt_d·0
Performance Characterization of SPEC CPU 2026 on AMD EPYC 9755 "Zen 5" Processorarxiv.org·2 pts·matt_d·0
Lost Bytes at the Crossroads Between User- and Kernel-Level Memory Allocation [pdf]ibr.cs.tu-bs.de·1 pts·matt_d·0