Entities · Models
Hugging Face
27 articles tagged with this entity.
-
The Blind Curator: How a Biased Judge Silently Disables Skill Retirement in Self-Evolving Agents
-
When Does In-Context Search Help? A Sampling-Complexity Theory of Reflection-Driven Reasoning
-
When do prophets profit in prediction markets?
-
Predicting the Emergence of Induction Heads in Language Model Pretraining
-
SynSFX: Multi-Model Sound Effects Synthesis Dataset for Deepfake Detection and Evaluation
-
Collaborative Disagreement Resolution for Scalable Oversight
-
Vertigo Vertigo: Reconstructing a Cinematic Ideal through its Predictive AI Double
-
Generating consensus and dissent on massive discussion platforms with a semantic-vector model
-
BayesBench: Evaluating LLM Belief Trajectories Under Multi-Turn Evidence Accumulation
-
SpreadsheetBench 2: Evaluating Agents on End-to-End Business Spreadsheet Workflows
-
Epiphany-Aware KV Cache Eviction Without the Attention Matrix
-
PhoneBuddy: Training Open Models for Agentic Phone Use
-
AGORA: An Archive-Grounded Benchmark for Agentic Workplace Document Reasoning
-
BenchX: Benchmarking AI Models for Cancer Detection and Localization with Demographic and Protocol Biases
-
Metis: Bridging Text and Code Memory for Self-Evolving Agents
-
Emergent Alignment
-
CareTransition-Audit: A Benchmark to Audit Discharge Summaries for Efficient Care Transitions
-
Attribute Inference from Interactive Targeted Ads
-
Revealing Artifacts via Noise Amplification: A Novel Perspective for AI-Generated Video Detection
-
Vernier: Probing Representational Misalignment Behind Lexical Gaps in Causal Reasoning
-
Automatic identification of diagnosis from hospital discharge letters via weakly supervised Natural Language Processing
-
Sycophancy as a Multilingual Alignment Failure: How Safety Degrades Across Languages, Topics, and Models
-
Hummus: A Dataset of Humorous Multimodal Metaphor Use
-
Taming Perception Jitter: Uncertainty-Aware LiDAR Object Detection for Reliable Motion Classification
-
Automated Framework to Evaluate and Harden LLM System Instructions against Encoding Attacks
-
Rewrite to Translate, Translate to Reward: Reinforcement Learning for Source Rewriting in Machine Translation
-
Depth over Fidelity in Fixed-Budget Noisy Evolution Strategies