跳到正文
//

人工智能

6 篇 · 资讯流
  1. arXiv cs.AI(人工智能)

    Large Language Models as Falsifiers for Cyber-Physical Systems

    Falsification searches for counterexamples to formal specifications in cyber-physical systems (CPS). With specifications

    原文 ↗
  2. arXiv cs.AI(人工智能)

    Language-model groups overstate consensus when replaying human deliberation on a reasoning task

    Full-consensus rates are often treated as indicators of collective cognition, yet depend on how participation and final

    原文 ↗
  3. arXiv cs.CL(自然语言)

    KoNeoBench: A Curated Evaluation Dataset for LLM Understanding of Korean Neologisms

    Large language models (LLMs) are typically evaluated on static benchmarks, even though natural language constantly evolv

    原文 ↗
  4. arXiv cs.DB(数据库)

    DualSQL: Text-to-SQL with Multi-Agent Reinforcement Learning

    State-of-the-art Text-to-SQL systems are typically multi-agent pipelines centered around two fundamental tasks: schema l

    原文 ↗
  5. arXiv cs.DB(数据库)

    The Stochastic Deputy: Structural Tenant Isolation for Tool-Using LLM Agents

    Multi-tenant tools commonly accept a tenant identifier and validate it against the caller's entitlement. For a large lan

    原文 ↗
  6. arXiv cs.DB(数据库)

    Natural Language Knowledge Graph Query Execution: Leveraging Controlled Semantics in the LLM Context Window

    Large Language Model (LLM) applications often transfer domain concepts into the model's context informally, through prom

    原文 ↗