Entities · Products
LLM
36 articles tagged with this entity.
-
Is Your NPU Ready for LLMs? Dissecting the Hidden Efficiency Bottlenecks in Mobile LLM Inference
-
Chasing Moving Targets with Online Self-Play Reinforcement Learning for Safer Language Models
-
Hyperloop Transformers
-
Gravity-Awareness: Deep Learning Models and LLM Simulation of Human Awareness in Altered Gravity
-
CausalMix: Data Mixture as Causal Inference for Language Model Training
-
How Human Feedback Shapes AI-generated Community Notes
-
SA-VLA: State-aware tokenizer for improving Vision-Language-Action Models' performance
-
Self-Stigma Is Not a Monolith, but Generic Empathy Is: Persona-Conditioned LLM Support for People Who Use Drugs
-
Contagion Networks: Evaluator Preference Propagation in Multi-Agent LLM Systems
-
ToE: A Hierarchical and Explainable Claim Verification Framework with Dynamic Multi-source Evidence Retrieval and Aggregation
-
HiLSVA: Design and Evaluation of a Human-in-the-Loop Agentic System for Scientific Visualization
-
How Do Tool-Augmented LLM Agents Perform on Real-World Energy Analytics Tasks?
-
A Deterministic Control Plane for LLM Coding Agents
-
A Probabilistic Framework for LLM-Based Model Discovery
-
AI translation of literary texts is "fine", but readers still prefer human translations
-
From 50K to 8.2 Million in 24 Hours: Vozinha's Algorithmic Consecration and the Multilingual Making of World Cup Visibility
-
AutoPass: Evidence-Guided LLM Agents for Compiler Performance Tuning
-
GLARE: A Natural Language Interface for Querying Global Explanations
-
Code-Augur: Agentic Vulnerability Detection via Specification Inference
-
SkillJect: Effectively Automating Skill-Based Prompt Injection for Skill-Enabled Agents
-
Whose hotel does the AI recommend? An algorithm audit of reputation signals in LLM-assisted hotel selection
-
IMPACTeen: Intentions, Manipulation, Persuasion, Annotations, and Consequences in Teen Communication Dataset
-
Mojo: A Promising Tool for Scalable Financial AI Efficiency
-
Building Customer Support AI Agents at 100M-User Scale: An Evaluation-Driven Framework
-
MiroBench: Benchmarking Realism in Agentic Simulation of Real-world Discussions
-
When Errors Become Narratives: A Longitudinal Taxonomy of Silent Failures in a Production LLM Agent Runtime
-
VeriGeo: Controllable Geometry Question Generation with Numerical and Analytical Verification
-
WISE: A Long-Horizon Agent in Minecraft with Why-Which Reasoning
-
EPIG: Emotion-Based Prompting for Personalised Image Generation
-
Brick: Spatial Capability Routing for the Mixture-of-Models (MoM) Paradigm
-
MetaPlate: Counterfactual-Guided RAG-LLM Tool for Personalized Food Recommendation and Hyperglycemia Prevention
-
Does Persona Make LLMs K-pop Fans? A Pilot Study of LLM-Based Online Concert Audience Agents
-
AliyunConsoleAgent: Training Web Agents in Real-World Cloud Environments via Distillation and Reinforcement Learning
-
Training for Technology: Adoption and Productive Use of Generative AI in Legal Analysis
-
A Dynamic Self-Evolving Extraction System
-
Running Python code in a sandbox with MicroPython and WASM