Working with the American Psychological Association on youth mental health and AIOpenAI and the American Psychological Association advance evidence-based guidance, resources, and safeguards for responsible AI use and youth mental health.·AI/AI 应用·openai.com ↗openai.com ↗
Architectural Implications of Agentic AI WorkflowsarXiv:2608.04458v1 Announce Type: new Abstract: Agentic AI is emerging in datacenters, but its architectural implications remain unexplored. We organize agentic workflows in a taxonomy and present its first architectural·AI/AI 应用, Agent·arxiv.org ↗arxiv.org ↗
Improving Auto-Design of Neural PDE Solvers with a Domain-Specific LanguagearXiv:2608.04384v1 Announce Type: new Abstract: Neural PDE solver auto-design is fundamentally a search-space representation problem. In the space of unrestricted Python programs, valid solvers form an extremely sparse s·AI/航天·arxiv.org ↗arxiv.org ↗
NeuMoSync: End-to-End Neuromodulatory Control for Plasticity and Adaptability in Continual LearningarXiv:2608.04358v1 Announce Type: new Abstract: Continual learning (CL) requires models to learn tasks sequentially, yet deep neural networks often suffer from plasticity loss and poor knowledge transfer, which can imped·AI/深度学习·arxiv.org ↗arxiv.org ↗
The RAIL Principles for Neurosymbolic AI: Reasoning, Assurances, Interfacing and LearningarXiv:2608.04285v1 Announce Type: new Abstract: Neurosymbolic AI systems that integrate machine learning and symbolic reasoning are rapidly gaining attention. They complement the data-intensive statistical approaches of·AI/AI 应用, 推理优化·arxiv.org ↗arxiv.org ↗
MatrAIx: Simulating the World with 8.3 Billion Persona AgentsarXiv:2608.04205v1 Announce Type: new Abstract: Human evaluation of AI systems and digital products is costly, slow, and difficult to scale. Offline evaluations are more scalable but often abstract away human diversity a·AI/AI 应用, 智能体·arxiv.org ↗arxiv.org ↗
Adversarially Robust Abductive Fusion of Pre-trained Transformer-based Perception ModelsarXiv:2608.04190v1 Announce Type: new Abstract: Deploying pre-trained perception models in novel environments degrades their accuracy under distributional shift, and assembling them alone does not recover it: combiners s·AI/轻小说·arxiv.org ↗arxiv.org ↗
FinPerMA: A Theory-Informed, Event-Grounded Personalized-Memory Benchmark for LLM AgentsarXiv:2608.04095v1 Announce Type: new Abstract: Large language model (LLM) agents are increasingly used as personalized assistants in high-stakes domains such as financial advising, yet it remains unclear whether they ca·AI/LLM, 半导体, 智能体·arxiv.org ↗arxiv.org ↗
FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Professional DeliverablesarXiv:2608.04077v1 Announce Type: new Abstract: Evaluating financial AI agents requires criteria aligned with real professional work. Existing rubric methods typically derive criteria from task prompts or model outputs,·AI/AI 应用, 智能体·arxiv.org ↗arxiv.org ↗
The LLM Proposes, the Executive Disposes: A Self-Verifying Agent Instrument that Dissociates Commitment Drift from Binding Drift in Long-Horizon AgentsarXiv:2608.04066v1 Announce Type: new Abstract: How do you verify a long-horizon agent when its own state and self-reports are exactly what you cannot trust? We present an agent instrument built so that verification is s·AI/Agent, LLM, 智能体·arxiv.org ↗arxiv.org ↗
A Long-Run Persistence Theory for AI Systems under the Redundancy-Adjusted Artificial Age ScorearXiv:2608.04012v1 Announce Type: new Abstract: Artificial intelligence systems are increasingly expected to operate over repeated cycles of interaction, adaptation, and update rather than through isolated one-shot outpu·AI/AI 应用·arxiv.org ↗arxiv.org ↗
阿里达摩院开放 15 项芯片与 AI 研究课题,启动“阿里星”顶尖人才招募IT之家 8 月 6 日消息,阿里达摩院今日宣布了面向 2027 届毕业生 的“阿里星”顶尖人才招募培养计划。 达摩院今年已开放 15 项研究课题 ,涉及 AI 芯片、新型 CPU 架构、医疗多模态智能体、 AGI 决策等众多前沿探索。IT之家付具体课题如下: 芯片方向 AI 芯片编程模型设计实现 为自研 AI 芯片打造全新编程语言和编译器工具链,不是简单的 " 移植 LLVM 后端 ",而是覆盖 DSL 设计、编译器前端、中端优化框架·AI/AI 应用, 多模态, 智能体, +1·ithome.com ↗ithome.com ↗
From asking to doing: How the world is putting ChatGPT to workNew OpenAI Signals data shows how people use ChatGPT worldwide, with country-level insights on adoption, usage trends, and evolving behavior.·AI/AI 应用·openai.com ↗openai.com ↗
消息称美国前沿人工智能模型网络安全审查机制暂不涉及开放权重模型IT之家 8 月 5 日消息,综合《纽约时报》《华尔街日报》《Axios》、彭博社等外媒报道,美国政府在当地时间本周二与头部人工智能企业举行的闭门会议中表示, 其今年 6 月宣布的前沿模型网络安全审查机制当前仅涉及闭源模型 , 开放权重模型至少暂时不会受到影响 。 知情人士透露,美国政府在会议中将需自愿接受审查的模型定义为具有最先进功能和国家安全风险的闭源模型, 该机制的任何内容都不应被解释为限制已发布的开源模型 。 图源:Pexels·AI/AI 应用·ithome.com ↗ithome.com ↗
阿里云上线 One Key MCP 服务:兼容 Qoder、Codex 等,可一键调用多家 MCP 服务IT之家 8 月 5 日消息,阿里云今日宣布上线 One Key MCP 服务。开发者可通过统一的阿里云百炼 API Key 调用所有生态伙伴 MCP 服务, 兼容 Qoder、Codex、Claude Code、Cursor 等主流 Coding Agent ,简化多 MCP 服务的接入、鉴权与计费管理流程。 此前开发者接入不同 MCP 服务需分别申请凭证、配置鉴权并完成联调;服务数量增加后,接入和维护成本也随之上升。 One Key·AI/Agent, 大模型, 智能体·ithome.com ↗ithome.com ↗
卡帕西提出测试 AI 能力新思路:100 万 Tokens 额度构建《指环王》3D 交互场景IT之家 8 月 5 日消息,OpenAI 联合创始人安德烈 · 卡帕西(Andrej Karpathy)于 8 月 2 日在 X 平台发布推文,提出一项 AI 基准测试思路: 向模型提供《指环王》开篇首段文字,再要求模型创建对应的 3D 世界。 在测试 AI 模型时,此前业内习惯使用生成“鹈鹕骑车图”的 SVG 图片方式,用来评估大语言模型在空间推理、几何逻辑和复杂代码生成方面的真实水平。 不过卡帕西指出目前模型能力不断提升,等简单任·AI/AI 应用, 大模型·ithome.com ↗ithome.com ↗
Neurosymbolic Reasoning with Incremental Knowledge for Sample Efficient Hierarchical Reinforcement LearningarXiv:2608.02993v1 Announce Type: new Abstract: (Flat) Reinforcement Learning (RL) agents face significant challenges in environments with sparse rewards that require long-horizon reasoning. A compelling approach to impr·AI/推理优化, 智能体·arxiv.org ↗arxiv.org ↗
On the missing data layer and a potential solutionarXiv:2608.02949v1 Announce Type: new Abstract: Latin America is missing two foundational layers of AI infrastructure: the dataset layer and the benchmark layer. This paper targets the dataset layer. The dataset layer fa·AI/AI 应用·arxiv.org ↗arxiv.org ↗