Skip to main content
autorenew
close
Yunduiyou AI
Write once, run everywhere.
Home
Blog
Archive
Projects
About
translate
🇺🇸
English
简体中文
search
⌘K
Archive
2026
Sep 2
OpenAgentFlow: Enabling System-Wide Safety Boundaries for Heterogeneous AI Agent Fleets
Sep 2
Incremental Risk Assessment of Progressive Elder Financial Scams via Instruction-Tuned Small Language Models
Sep 2
UI-Venus-2 Technical Report
Sep 2
I-CARE: Analysis of interference-related phenomena in a controllable, diverse and representative unlearning setting for text-to-image models
Sep 2
HyperWorld: Hypergraph-Structured State Serialization Improves Learned Textual World Models
Sep 2
When Prediction Error Is Not Enough: Evaluating Nuisance-Function Prediction for Causal Estimation
Sep 2
Discrete-Time MDP Modeling for Multi-Item Capacitated Lot Sizing with Stochastic Demand Timing
Sep 2
Long-Horizon State Tracking in LLMs: Executing MD5 through a Deep Sequence of Dependent Tool Calls
Sep 2
SCAFFOLD: A Large-Scale Structured Dataset of Computer Science Research Figures with Diagram QA and Chain-of-Thought Reasoning Traces
Sep 2
EULER: Exploring Underused Links with Evidence-Checked Return for Multi-Agent Mathematical Discovery
Sep 1
Expert-validated STEM QA
Sep 1
Paper Pilot: A Human-in-the-Loop Expert System for Evidence-Traceable Scientific Manuscript Generation in Applied Sciences
Sep 1
Introducing @huggingface/kernels: 200+ WebGPU Kernels for Local AI
Sep 1
BenchMIRT: What are LLM benchmarks actually measuring?
Sep 1
Statutory AI: Aligning Large Language Models With Legal Norms
Sep 1
The Signal in the Noise: An Auditable Reliability Layer for Biomedical Text Classification
Sep 1
From Question-First to Analyst-First: Domain-Expert Skills and Verified Knowledge Compilation for Proactive Enterprise Analytics
Sep 1
The Race between Agentic AI Capabilities and Data Quality Control in Online Surveys
Sep 1
DS-Lighting: Making Agent Harnesses Explicit for Data-Science Automation
Sep 1
CDPR: Counterfactual Advantage-based Credit Assignment for Cost-Aware Sequential Medical Diagnosis
Sep 1
SHAPE of Chain-of-Thought in Math Reasoning
Sep 1
A collective capability boundary in frontier large language models on guideline-conformant and case-specific oncology decision-making
Aug 28
The Open ASR Leaderboard Adds Its First Global South Language
Aug 26
Training and Finetuning Multi-Vector Embedding Models with Sentence Transformers
Aug 25
Wire It, Run It, Deploy It: AI Workflows in Gradio
Aug 25
Granite 4.2 LLMs: How They're Built
Aug 25
Quantization-Aware Healing: a compressed, 4-bit model that outperforms its full-precision original
Aug 21
Measuring benchmark optimization in speech recognition
Aug 21
How Hugging Face Inference Endpoints, Jobs, and Buckets Power Search on Papers with Code
Aug 20
Up to 3.2x Faster Inference with LFM2.5-DSpark
2025
Jan 1
Welcome to Your Blog
Search posts…
⌘K or Ctrl+K to open search
Search by
Algolia
No results.