//
人工智能
6 篇 · 资讯流Large Language Models as Falsifiers for Cyber-Physical Systems
Falsification searches for counterexamples to formal specifications in cyber-physical systems (CPS). With specifications
原文 ↗Language-model groups overstate consensus when replaying human deliberation on a reasoning task
Full-consensus rates are often treated as indicators of collective cognition, yet depend on how participation and final
原文 ↗KoNeoBench: A Curated Evaluation Dataset for LLM Understanding of Korean Neologisms
Large language models (LLMs) are typically evaluated on static benchmarks, even though natural language constantly evolv
原文 ↗DualSQL: Text-to-SQL with Multi-Agent Reinforcement Learning
State-of-the-art Text-to-SQL systems are typically multi-agent pipelines centered around two fundamental tasks: schema l
原文 ↗The Stochastic Deputy: Structural Tenant Isolation for Tool-Using LLM Agents
Multi-tenant tools commonly accept a tenant identifier and validate it against the caller's entitlement. For a large lan
原文 ↗Natural Language Knowledge Graph Query Execution: Leveraging Controlled Semantics in the LLM Context Window
Large Language Model (LLM) applications often transfer domain concepts into the model's context informally, through prom
原文 ↗