Orca-Bench: How Ready Are Language Model Agents for Oncall?

Anthropomorphism in Children's Interactions with LLM Chatbots

How real are real numbers? (2004)

Time-Frequency Consistency Learning for Robust Speech Deepfake Detection

Honey Bee Colony Monitoring via Audio IoT Sensors, Tensorgrams and RNNs

The Burau representation of the braid group is faithful for n = 4

A Measurement Study on the Adoption of Pledges and Unveils in the OpenBSD Operating System

SEAM-V: A Hybrid-Decoupled RISC-V Vector Processor

Agent Security Is a Systems Problem

Ring-Zero: Scaling Zero RL to a Trillion Parameters for Emergent Reasoning

ESBMC-Arduino: Closing the Deployment Gap for Formal Verification

Qubes OS Security in the Public Record

Mathematics of Data Science

Greedy is optimal for single-pass semi-streaming matching

Co-evolution of self-replication and function in a digital primordial soup

Kani: A Model Checker for Rust

Investigating idiosyncrasies in AI fiction

Billions of Sketches Reveal Hidden Cultural Variation in Human Concepts

Rzk: A Proof Assistant for Synthetic ∞-Categories

Coding agents think ahead of time

The age of the Universe from a large sample of the oldest Galactic stars

Can LLMs Perform Deep Technical Comprehension of Computer Architecture Papers

Fleet: Hierarchical Task-Based Abstraction for Megakernels on Multi-Die GPUs

Robust Secret Storage in Networks

Prismata: Confining cross-site prompt injection in web agents

A Study of Microsoft's Early 2026 Rollout of Claude Code and GitHub Copilot CLI

Exploiting LLM Agent Supply Chains via Payload-Less Skills

Taxing Artificial Intelligence

Automation Without Understanding

Protocol Prying: Vulnerability Research in AirDrop and Quick Share

More →