微信内可能无法直接打开本站。请点右上角 ··· → 在浏览器打开 ,或复制链接后用系统浏览器访问。
综合
官方
企业
汇聚
当前信源:arXiv cs.AI
· 清除信源筛选
1 / 18
InfraBench: Evaluating Infrastructure Agents Across Layers, Lifecycle, and Risk
arXiv:2608.11234v1 Announce Type: new Abstract: Managing modern computing infrastructure has become a steadily harder problem due to the ever-increasing complexity.…
Recent advances in AI agents create a timely opportunity to automate i…
We present InfraBench, a benchmark suite for evaluating AI agents on r…
Experiments with 15 agent-model configurations show that even the stro…
RSS 官方收录 · 可信分层展示
LinearKV: One Cached State Suffices for Position-Independent Caching in Hybrid LLMs
arXiv:2608.11231v1 Announce Type: new Abstract: LLM serving is increasingly accelerated by position-independent caching (PIC).…
Existing PIC methods, however, are built for full-attention models, wh…
Hybrid LLMs break these primitives---they replace most attention layer…
This raises a natural question: can PIC benefit hybrid models, and wha…
RSS 官方收录 · 可信分层展示
The Edge-based Contiguous p-median Problem with Connections to Logistics Districting
arXiv:2608.…
11230v1 Announce Type: new Abstract: This paper introduces the edge-ba…
Two binary programming models are introduced, both of which incorporat…
The first model requires an exponential number of cut set-based constr…
RSS 官方收录 · 可信分层展示
Synchronizing Beliefs with Second-Order Theory-of-Mind in Human-Autonomy Teams (Extended Version)
arXiv:2608.…
11229v1 Announce Type: new Abstract: Comparative feedback, asking peop…
Preference-based reward learning typically casts the human teacher as …
We argue this forfeits the teacher's defining advantage: knowledge of …
RSS 官方收录 · 可信分层展示
Forecasting Side Effects of Activation Steering
arXiv:2608.…
11227v1 Announce Type: new Abstract: Activation steering modifies a la…
While effective, steering often produces unintended side effects on ot…
We therefore ask: can these side effects be forecasted before steering…
RSS 官方收录 · 可信分层展示
Cutting AI Datacenter Energy with Reinforcement Learning: Measured Power Control of LLM Training from One GPU to the Fleet
arXiv:2608.…
11226v1 Announce Type: new Abstract: Reinforcement-learning post-train…
We instrument GRPO training with half-second power telemetry at 7B, 14…
Against the full 500-step 7B trace, the controller cuts power-limit vi…
RSS 官方收录 · 可信分层展示
Identity from the Outside: A Conceptual Framework and Research Program for AI Personality Clones
arXiv:2608.11225v1 Announce Type: new Abstract: AI "personality clones" force a re-examination of personal identity in operational terms.…
Setting aside the hard problem of consciousness, we approach identity …
We distinguish three criteria that "identity" conflates: fidelity to a…
We propose a six-term factorization of observed identity (substrate, d…
RSS 官方收录 · 可信分层展示
Harnessing agent memory to build lifelong AI partners for materials scientists
arXiv:2608.…
11224v1 Announce Type: new Abstract: Materials research advances throu…
This experience is essential for reproducibility and knowledge transfe…
Here we argue that a lifelong AI partner for materials science can be …
RSS 官方收录 · 可信分层展示
A Conceptual Framework for Refining Influence Knowledge from Simulation Evidence in Cyber-Physical Systems
arXiv:2608.…
11221v1 Announce Type: new Abstract: Cyber-physical systems (CPS) are …
The behaviour of these systems emerges from the interaction between th…
Simulation and co-simulation have become essential approaches for anal…
RSS 官方收录 · 可信分层展示
LLMs in Process Diagram Engineering: From Optimal PFDs to Validated P&IDs
arXiv:2608.…
11220v1 Announce Type: new Abstract: Nowadays, the creation of a proce…
Applying artificial intelligence in the task could potentially lead no…
This research presents P&ID Pilot - a practical end-to-end AI pipeline…
RSS 官方收录 · 可信分层展示
From Monolithic to Modular: Segment-level Automatic Prompt Optimization
arXiv:2608.11219v1 Announce Type: new Abstract: Automatic Prompt Optimization (APO) often rewrites prompts monolithically, which can improve one behavior while degrading others.…
We present SAPO, a segment-level APO method that decomposes prompts in…
The optimization loop uses one LLM with static meta-prompts and struct…
We describe a train/validation protocol and a two-stage generation pro…
RSS 官方收录 · 可信分层展示
MaSRead: Content-Addressed Reading of Replicated Latent Stores
arXiv:2608.11218v1 Announce Type: new Abstract: Independent agents that reason in latent space can share computed state as key-value cache fragments rather than text.…
Merged by a conflict-free replicated data type, these fragments form a…
Yet a later query, unknown at encode time, cannot reliably read the me…
MaSRead addresses the read to content.
RSS 官方收录 · 可信分层展示
AutoWorldModel-Bench: A State-Centric Benchmark for Automated World-Model Research
arXiv:2608.…
11216v1 Announce Type: new Abstract: World modeling is an unsettled fi…
This makes it an ideal testbed for AI coding agents acting as autonomo…
We introduce AutoWorldModel-Bench, a closed-loop benchmark in which fr…
RSS 官方收录 · 可信分层展示
Poor Man's Agentic Modeling: Simulating Large LLM-Agent Societies on a Laptop
arXiv:2608.…
11215v1 Announce Type: new Abstract: Simulating societies of many larg…
We turn a statistical-physics observation into a method: replace each …
Whether this works is decided before the simulation runs, chiefly by w…
RSS 官方收录 · 可信分层展示
Detecting a Route Flip Is Easier Than Knowing Whether to Fix It: Causal Route-Mediated Damage in Quantized Mixture-of-Experts
arXiv:2608.…
11212v1 Announce Type: new Abstract: Top-k Mixture-of-Experts (MoE) ro…
This paper proposes no new mitigation; it supplies a causal apparatus,…
A four-run apparatus prices the route-mediated fraction (RMF) of quant…
RSS 官方收录 · 可信分层展示
A Forced-Structure Reduction and Verifiable Bounds for Conway's 99-Graph
arXiv:2608.11211v1 Announce Type: new Abstract: Conway's 99-graph problem asks whether a strongly regular graph with parameters $\mathrm{srg}(99,14,1,2)$ exists.…
We report a systematic, fully reproducible attack by an autonomous AI …
Our verifiable contributions are: (1) an exhaustive proof that no circ…
0\%$ of the constraints ($33$ of $49$ difference-classes), with the sa…
RSS 官方收录 · 可信分层展示
Distribird: Literature-Informed Prior Distribution Design for Bayesian Model Calibration
arXiv:2608.11210v1 Announce Type: new Abstract: Bayesian calibration of process-based models requires a prior distribution for each model parameter.…
Despite decades of methodological work, researchers almost always fall…
The main reason is that building informative priors from scientific li…
We present \textbf{Distribird}, an agentic web application that automa…
RSS 官方收录 · 可信分层展示
Dynamic Governance of Multi-LLM Agent Systems for Collaborative Conversational Outcomes
arXiv:2608.…
11207v1 Announce Type: new Abstract: When two LLM agents with structur…
This paper asks whether a control-theoretic governance layer can subst…
The Experience Orchestrator (EO) addresses this in a simulated financi…
RSS 官方收录 · 可信分层展示
上滑下一条
上滑 · j/k · m/u · h 隐藏 · a 稍后 · o 原文 · e 详情 · t 今日 · i 模式 · f 搜索