Loading...

Tag trends are in beta. Feedback? Thoughts? Email me at [email protected]

Quantum information spreading via higher-order operator correlators

ArXiv receives multiyear commitments to support it as an independent nonprofit

Harm Laundering in GPT Models: Gender Discrimination Transformed Rather Than

LLM Judges Verify Presence, Not Absence: Omission Blindness in AI Clinical Notes

DeepSeek Elastic Compute (DSec): Sandbox Infrastructure for Effective Agentic Training at Scale

The Implications of Linguistic Illegibility for LLM Security

LLMs as a Cognitive Virus

Show HN: Training a model to identify AI web content from structure alone

How good are frontier models at physics?

Dream-RSI: Recursive Self-Improvement through Evolving Worlds

Trusting-Trust Attack against an Entire Linux Distribution

Compiler-style optimization for drawing via Skia

An empirical study of harness design for coding agents

Accurate Models of AMD Matrix Cores

Infinite-Parameter LLMs: Generating and Adapting Weights from Live Data

Thinking with Looped Flows

Cache-to-Cache: Direct Semantic Communication Between LLMs (2025)

Reflections on Trusting Trust, Revisited: Poisoning Self-Modifying AI Coding

The Malicious Use of Artificial Intelligence

Bye Bye Perspective API: Lessons for Measurement Infrastructure in NLP, CSS and LLM Evaluation

MISRust: Mapping MISRA-C++ Coding Guidelines to the Rust Programming Language

Not In My Git Yard: Catching Backdoors at Commit and Release Time

The Pain Axis: LLMs Represent Self-Directed Harm and Act to Relieve It

Randomized query complexity can beat certificate complexity

GRP-Obliteration: Unaligning LLMs with a Single Unlabeled Prompt

Will there be a 7G?

C*: Unifying Programming and Verification in C (2025)

The k-server conjecture is true

Breaking the 1.58-bit Barrier for Ternary LLMs

RoofLang: Enabling AI-Driven Architecting of LLM Inference Systems

More →