Research
Academic papers and industry research on AI systems and human-centred approaches.
Last updated: 1 August 2026
Anthropic
Emotion Concepts and Their Function in a Large Language Model
April 2026
Anthropic
Automated Alignment Researchers: Using Large Language Models to Scale Scalable Oversight
April 2026
arXiv
A Systematic Survey of Prompt Engineering in Large Language Models
February 2024
arXiv
AuditBench: Evaluating Alignment Auditing Techniques on Models with Hidden Behaviors
February 2026
arXiv
The Collaboration Gap in Human-AI Work
April 2026
arXiv
Scaling Human-AI Coding Collaboration Requires a Governable Consensus Layer
April 2026
arXiv
The Instrumental Dissolution of Typing: Why AI Challenges the Keyboard Era in Knowledge Work
April 2026
arXiv
EVA-Bench: A New End-to-end Framework for Evaluating Voice Agents
2026-05-13
arXiv
History Anchors: How Prior Behavior Steers LLM Decisions Toward Unsafe Actions
2026-05-13
arXiv
Improving Reproducibility in Evaluation through Multi-Level Annotator Modeling
2026-05-13
arXiv
Negation Neglect: When models fail to learn negations in training
2026-05-13
Brookings
The Effects of AI on Firms and Workers
July 2025
Google DeepMind
SIMA 2: Agents for 3D Virtual Worlds
November 2025
Harvard / BCG
Navigating the Jagged Technological Frontier: Field Experimental Evidence of the Effects of AI on Knowledge Worker Productivity and Quality
September 2023
OECD
The Impact of Artificial Intelligence on Productivity, Distribution and Growth
April 2024
OpenAI
PaperBench: Evaluating AI's Ability to Replicate AI Research
March 2025
Stanford HAI
AI Index Report 2026
April 2026