Entities · Products
iPhone 16
110 articles tagged with this entity.
-
Cost-Effective Agent Harnesses for Abstract Reasoning and Generalization on ARC-AGI-1
-
Do LLM-Generated Skills Make Better AI Data Scientists? A Component Ablation Across Data-Science Workflows
-
When Agents Go Rogue: Activation-Based Detection of Malicious Behaviors in Multi-Agent Systems
-
LLMs Silently Correct African American English: Auditing and Mitigating Dialect Bias via Activation Steering
-
EgoPolice: A Benchmark for Egocentric Video Understanding in High-Stakes Police Body-Worn Camera Footage
-
MLLM-LLaVA-FL: Multimodal Large Language Model Assisted Federated Learning
-
InfluMatch: Frontier-Quality KOL Search at 4B-Model Cost
-
Native-speed vLLM transformers modeling backend
-
sqlite-utils 4.0, now with database schema migrations
-
Shut Those Laptops! Anthropic Puts Its Claude Cowork Agent on Your Phone
-
Solos debuts an even lighter version of its camera-less smart glasses
-
Shark ChillPill 3-in-1 fan review: the handheld fan I’d pack for every trip – at a price that’ll make you sweat
-
Memory-Orchestrated Semantic System (MOSS): An Auditable Agentic Memory Architecture
-
Predicting Drafted Deck Strength for "Magic: the Gathering"
-
REDDIT: Correcting Model-Generated Timestamp Drift in ASR without Forgetting via Replay-Based Distribution Editing
-
Track the Noise, Move the World:3D-Grounded Motion-Consistent Noise for Controllable Video Generation
-
An automated method of identifying incorrectly labelled images based on the sequences of loss functions of deep learning networks
-
S-EMBER: A Large-Scale Benchmark for Streaming Egocentric Memory Retrieval
-
A 10,000-Year Global Stochastic Tropical Cyclone Catalog with Wind-Dependent Track Transitions (WHITS)
-
Federated Learning for Object Detection: Enabling Collaborative Drone Learning Without Centralizing Data
-
Overloading Large Vision-Language Models for Jailbreaking
-
Reflective Dialogue or Prompt Refinement? Effects of Tutor Scaffolding on Students' Independent LLM Use for Programming
-
Language models guide symbolic equation discovery by controlling search
-
Microsoft cuts 4,800 jobs as it revamps Xbox in latest wave of mass layoffs
-
Conditional Co-Ablation: Recovering Self-Repair Backups in Transformer Circuits
-
SA-HGNN: Sample-Adaptive Hyperbolic Graph Neural Network for EEG-Based Depression Recognition
-
AIriskEval-edu: New Dataset for Risk Assessment in AI-mediated K-12 Educational Explanations
-
Bounded Morality: Defining the Space of Moral Computation
-
Theory of Mind and Persuasion Beyond Conversation: Assessing the Capacity of LLMs to Induce Belief States via Planning and Action
-
[AINews] Sonnet 5 today, and Fable 5 tomorrow
-
Forward Deployed Engineers and the future of software engineering
-
HERO: Improving the Reliability and Sensitivity of Generative Model Evaluation Using Historical Data
-
Learning Efficient 4D Gaussian Representations from Monocular Videos with Flow Splatting
-
CaresAI at CT-DEB26: Detecting Dosing Errors In Clinical Trials Using Domain-Specific Transformer Embeddings and Classification Models
-
PGE-SAM: Prompt-Guided Feature Enhancement for Interactive Segmentation under Degradation
-
AMR: Adaptive Modality Routing for Multimodal Polyglot Speaker Identification
-
Featuring Every Eval Ever Results on Hugging Face Model Pages
-
Ring Video Doorbell Pro review: night and day better with new 4K camera
-
MVPruner: Dynamic Token Pruning for Accelerating Multi-view Vision-Language Models in Autonomous Driving
-
EMOSH: Expressive Motion and Shape Disentanglement for Human Animation
-
Home3D 1.0: A High-Fidelity Image-to-3D Asset Generation System for Interior Design
-
Batch-Invariant Spectral Intelligence for Robust and Explainable Insect Authentication
-
In-Context Model Predictive Generation: Open-Vocabulary Motion Synthesis from Language Models to Physics
-
LISA: Likelihood Score Alignment for Visual-condition Controllable Generation
-
A-Evolve-Training: Autonomous Post-Training of a 30B Model
-
Assert, don't describe: Linguistic features that shift LLM reasoning about animal welfare
-
WatchAct: A Benchmark for Behavior-Grounded Robot Manipulation
-
SAM2Matting: Generalized Image and Video Matting
-
Closing the Quality Gap in Low-Resource Text-to-Speech: LoRA Fine-Tuning of VoxCPM2 for Khmer and Korean
-
Dyson HushJet Mini Cool fan review: I’ve never tested a handheld fan this powerful – or this loud
-
AI Is Designing Radio Chips That Humans Couldn’t Even Imagine
-
The emergence of the web data infrastructure layer for AI
-
CORE-Bench: Fostering the Credibility of Published Research Through a Computational Reproducibility Agent Benchmark
-
EG-VQA: Benchmarking Verifiable Video Question Answering with Grounded Temporal Evidence
-
Cycle-Consistent Neural Explanation of Formal Verification Certificates
-
Nine Judges, Two Effective Votes: Correlated Errors Undermine LLM Evaluation Panels
-
Google DeepMind bets $75M on AI’s future in Hollywood with A24 deal
-
Hotter Than a Hot Tub: The 45°C Breakthrough to Cool AI’s Biggest Machines
-
28 Tips to Take Your ChatGPT Prompts to the Next Level
-
The Register Gap: A Meaning Intelligence Framework for Nigerian Public Discourse
-
What sentiment analysis can't see: Measuring whether customers were helped, and what went wrong, across 70,000 support conversations
-
LaViSA: A Language and Vision Structural Ambiguity Benchmark
-
Contour-Constrained Deformable Registration with Parameter Characterization for Head and Neck Surgical Guidance
-
Thermodynamic Signatures of Reasoning: Free-Energy and Spectral-Form-Factor Diagnostics for Hallucination Detection in Large Language Models
-
Amazon employees say they’re facing termination for backing data center limits
-
The Adventures of Elliot: The Millennium Tales review – a playable love letter to Zelda
-
The UK Will Scan Asylum-Seekers’ Faces for Age Checks—Despite Knowing the Tech Is Flawed
-
From Memorization to Creation: Evaluating the Cognitive Depth of LLM-Generated Educational Questions
-
RedactionBench
-
CAPRA: Scaling Feedback on Software Architecture Deliverables with a Multi-Agent LLM System
-
State of the blog, mid-2026
-
Public transit gains and spatially uneven travel demand changes after NYC congestion pricing
-
ActWorld: From Explorable to Interactive World Model via Action-Aware Memory
-
CogCanvas: A Benchmark for Evaluating Multi-Subject Reference-Based Image Generation
-
DRA-GRPO: Your GRPO Needs to Know Diverse Reasoning Paths for Mathematical Reasoning
-
FP8 is All You Need (Part 1): Debunking Hardware FP64 as the HPC Holy Grail (June 13th version)
-
FinBalance: A Multi-Document Accounting Reconciliation Benchmark
-
AI Pluralism and the Worlds It Misses
-
Orchestrated Reality: From Role-Play to Living, Playable Game Worlds -- LLM-Driven World Simulation as a Parameterized-Action POMDP
-
IndustryBench-MIPU: Benchmarking Multi-Image Attribute Value Extraction for Industrial Products
-
No Accidental Software Agent First Canonical Code for Human Code Entropy Reduction and 30 to 500 times Lower Frontier Model Requirements
-
High-Frequency Pricing at Scale for E-Commerce
-
HarnessX: A Composable, Adaptive, and Evolvable Agent Harness Foundry
-
My yard is dying, so I made an app for that
-
Person Identification from Contextual Motion
-
SalArt-VQA: Diagnosing Whether VLMs Understand Salient Artifacts in Generated Images
-
Iterating Toward Better Search: A Two-Agent Simulation Framework for Evaluating Agentic Search Architectures in E-Commerce
-
X-MADAM-RAG: Diagnosing and Handling Chinese-English Evidence Conflict in Retrieval-Augmented Generation
-
LLMs Can Better Capture Human Judgments--With the Right Prompts
-
SpaceX officially prices shares at $135 in the largest IPO ever
-
The best robot vacuums in the UK to keep your home clean and dust free, tested
-
Alignment Defends LLMs from Property Inference Attacks
-
Does Reasoning Preserve Alignment? On the Trustworthiness of Large Reasoning Models
-
Understanding and mitigating the risks of OpenClaw for non-technical users: A practical guide with Skill
-
Structure from Reasoning, Numbers from Search: On-Premise Open LLMs as Structural Priors for Coupled MIMO Controller Tuning
-
CollabSkill: Evaluating Human-Agent Collaboration On Real-World Tasks
-
Beyond Model Size: Probing the Gaps in Visual in-Context Learning by Training a Tiny Model
-
What it feels like to work with Mythos
-
What Do AI Standards Mean for Small and Medium Enterprises?
-
Integrating Deep Learning Demand Forecasting with Multi-Objective Optimization for Circular Coffee Supply Chains: A Data-Driven Framework for Cost, Emissions, and Freshness Management
-
Stress-testing medical large language models reveals latent safety pathology beyond benchmark accuracy
-
How Deep Are Deep GPs, Really? A Sharp Threshold and a Non-Gaussian Limit for Compositional GPs
-
FunctionEvolve: Structure-Guided Symbolic Regression with LLMs
-
What the Eyes See, the LLMs Miss: Exploiting Human Perception for Adversarial Text Attacks
-
Auditing Proprietary Alignment in Large Language Models: A Comparative Framework Without a Ground-Truth Standard
-
Apple’s WWDC AI demos looked more real after $250M false ad settlement
-
Bridging intent and execution in agentic systems
-
Summer Game Fest highlights: 34 new video games to look out for, from Alien Isolation to Crazy Taxi
-
The Masked Advantage: Uncovering Local-Language Access to Cultural Knowledge in LLMs
-
Perplexity Can Miss SAE Feature Damage Under Quantization