| 1 |
Automated Design of Inventory Policy with Large Language Models: An Exploratory Study |
提出基于大语言模型的自动化库存政策设计方法 |
large language model |
|
|
| 2 |
Answer-Distribution Trajectories: A Stochastic-Dynamics View of LLM Reasoning |
提出答案分布轨迹以解决LLM推理动态分析问题 |
chain-of-thought |
|
|
| 3 |
Vision: Data-Centric Anchoring for Robust and Interpretable Agentic AI |
提出数据中心锚定以解决代理AI的鲁棒性与可解释性问题 |
large language model |
|
|
| 4 |
Procedural Graphs: Self-Evolving Execution Structures for LLM Agents |
提出程序图以解决长时间规划中的行动选择问题 |
large language model |
|
|
| 5 |
MeClear: Cooperative Game-Theoretic Attribution and Risk-Aware Memory Clearance for Long-Horizon LLM Agents |
提出MeClear以解决长时间交互中的记忆管理问题 |
large language model |
|
|
| 6 |
SQLMorph: Query Mutation and Fine-Grained Metrics for Text-to-SQL Evaluation |
提出SQLMorph以解决Text-to-SQL评估中的复杂性问题 |
large language model |
|
|
| 7 |
It's All in the Way You Say It: The Role of Information Representation in LLM-Based Glycemic-Event Prediction |
基于提示的LLM方法提升糖尿病患者血糖事件预测准确性 |
large language model |
|
|
| 8 |
Benchmark Scores Are Pipeline-Dependent: A Reliability Audit of Cybersecurity LLM Benchmarks |
审计网络安全LLM基准,揭示评分依赖于评估管道 |
large language model |
|
|
| 9 |
Graph-Based Personalized Memory for LLM Agents: Representation, Evolution, Retrieval, and Evaluation |
提出图基个性化记忆以解决LLM代理的个性化问题 |
large language model |
|
|
| 10 |
The Unreliable Progress Bar: Can LLM Agents Reliably Report Task Progress Throughout Execution? |
评估大型语言模型在任务执行中报告进度的可靠性 |
large language model |
|
|
| 11 |
AgentGrad: Intervention-guided Prompt Optimization for Multi Agent Systems |
提出AgentGrad以解决多智能体系统中的提示优化问题 |
large language model |
|
|
| 12 |
FastE: Readout-Triggered Token Compression for LLM Embedding Inference |
提出FastE以解决LLM嵌入推理中的前缀冗余问题 |
multimodal |
|
|
| 13 |
Noise Adaptive Streaming Audio-Visual Speech Token Enhancement for Robust Full-Duplex Spoken Dialogue Models |
提出AV-STE以解决全双工对话系统在噪声环境下的鲁棒性问题 |
multimodal |
|
|
| 14 |
LEBGen: An LLM-Enhanced Bayesian Network Framework for Few-Shot Travel Survey Data Generation |
提出LEBGen框架以生成少量旅行调查数据 |
large language model |
|
|
| 15 |
MemForest: Efficient Agent Memory Management via EventTree Partitioning and Progressive Merging |
提出MemForest以解决代理记忆管理效率问题 |
multimodal |
✅ |
|
| 16 |
Evidence-Aligned Entity Verification for Hallucination Detection in Retrieval-Augmented Generation |
提出证据对齐实体验证方法以解决检索增强生成中的幻觉检测问题 |
large language model |
|
|
| 17 |
ACEA: An Adversarial Co-Evolution Arena for Head-to-Head Red-Team and Blue-Team LLM Testing |
提出ACEA以解决大型语言模型的红队与蓝队测试问题 |
large language model |
|
|
| 18 |
TTGBench: Benchmarking Topological Evolution and Semantic Drift in Text-attributed Temporal Graphs |
提出TTGBench以解决文本属性时间图中的语义漂移问题 |
large language model |
|
|
| 19 |
Does Deeper Reasoning Compromise Alignment? Revealing and Mitigating of Alignment Collapse in Large Reasoning Models |
提出Alignment Loss Rate以解决深度推理中的对齐崩溃问题 |
chain-of-thought |
|
|
| 20 |
Less Is Personal: Learning Minimal Sufficient User Profiles for Personalized Language Models |
提出ENOUGH方法以解决个性化语言模型的冗余用户记录问题 |
large language model |
|
|