When Compression Scores Cannot Decide: Information Boundaries for Group-Robust LLM PruningarXiv:2608.02940v1 Announce Type: new Abstract: A reproducible compression statistic can still select the wrong candidate. A dense pruning score with 0.906 split-half reliability predicted a 16.1% gain. Its selected endp·AI/LLM·arxiv.org ↗arxiv.org ↗
HyperAgent: Planning and Acting over Tool-Schema Hypergraphs for Tool-Use LLM AgentsarXiv:2608.02650v1 Announce Type: new Abstract: Large language model (LLM) agents increasingly rely on external tools to complete complex real-world tasks. However, reliable tool-use planning remains challenging due to t·AI/LLM, 智能体·arxiv.org ↗arxiv.org ↗
ISEE: Interactive Semantic Enrichment for Database FieldsarXiv:2608.02604v1 Announce Type: new Abstract: LLM-based agents are increasingly being deployed for data-related tasks, including data sense-making, exploration, and retrieval. However, their performance heavily depends·AI/LLM, 智能体·arxiv.org ↗arxiv.org ↗
Anthropic 任命库埃拉尔为首位全球事务官,负责应对 AI 政策IT之家 8 月 5 日消息,Anthropic 今天(8 月 5 日)发布公告, 宣布任命马里亚诺 - 弗洛伦蒂诺 · 库埃拉尔(Mariano-Florentino Cuéllar)为公司首任全球事务主管。 IT之家援引路透社报道,库埃拉尔将直接向公司联合创始人兼总裁丹妮拉 · 阿莫迪(Daniela Amodei)汇报工作,工作地点位于 Anthropic 的旧金山总部,领导公司在全球范围内的政策、战略国际参与和政府关系方面的工作·AI/AI 应用, 国际要闻, 政策·ithome.com ↗ithome.com ↗
Third-party cyber evaluations involving OpenAI modelsOpenAI explains recent third-party cybersecurity evaluation incidents and outlines new safeguards to strengthen AI model testing and evaluation.·AI/AI 应用·openai.com ↗openai.com ↗
LFM2.5 2.6B model competitive with 4x larger models77 points · 18 comments · by nateb2022·AI/大模型·huggingface.co ↗huggingface.co ↗
Is AI making us dumber? Maybe not. But our skills are at riskStudies suggest AI can weaken learning when it replaces effort, but tools that guide rather than answer may help people keep their skills.·AI/AI 应用·sciencenews.org ↗sciencenews.org ↗
中国“银河画卷”巡天证实星际分子云是超高能宇宙射线诞生的关键载体中新社青海德令哈8月4日电 (记者 孙睿)中国科学院紫金山天文台4日在第二届人工智能赋能的天文学开放科学会议上披露,位于青海德令哈的中国科学院紫金山天文台青海观测站13.7米口径毫米波射电望远镜完成“银河画卷(MWISP)”一期巡天后有跨界新发现,证实星际分子云是超高能宇宙射线诞生的关键载体,打通了分子天文与高能天文交叉研究全新赛道。·AI/AI 应用·chinanews.com.cn ↗chinanews.com.cn ↗
Linux 暂存区不欢迎 AI 生成补丁,安全修复需通过实际硬件测试验证IT之家 8 月 4 日消息,Linux 内核维护者、Linux 基金会研究员 Greg Kroah-Hartman 于当地时间周一宣布了一项新政策:针对 Linux 内核的暂存区(drivers/staging/),将不再接受由 AI 模型生成的补丁,但真实有效的安全漏洞修复除外。 Greg 解释称,近期 Linux 暂存区收到大量由 LLM 生成的补丁,因此有必要明确相关政策。他指出,暂存区的主要目的并不是维护成熟代码,而是作为新开·AI/AI 应用, LLM, 政策·ithome.com ↗ithome.com ↗
CoWoS封装产能告急,台积电进一步扩大外包产能英伟达的GPU订单已经把台积电的封装产线挤到极限,台积电不得不决定将AI芯片制造核心工席CoWos中的CoW(晶圆上芯片)封装产能进一步外包给日月光等封测厂商。CoWoS是台积电用于AI芯片的2.5D封装技术,AI处理器位于芯片中心,HBM(高带宽内存)环绕四周,通过中介层连接。(财联社)·AI/AI 应用, 半导体·36kr.com ↗36kr.com ↗
晶采观察丨“新新三样”全球爆单 折射中国经济向新向优向好在澳大利亚等地,中国研发的清洗机器人“飞檐走壁”,替代了“蜘蛛人”高空危险作业;来自上海的创新药胶囊被纳入全球权威治疗指南,为患者带来新希望……如今,由机器人、人工智能(AI)、创新药组成的“新新三样”在海外刷屏,成为中国外贸的新名片。不过,更值得关注的观察来自市场端。·AI/AI 应用·chinanews.com.cn ↗chinanews.com.cn ↗
Towards an AI for AfricaWestern-developed AI rides roughshod over Ubuntu values of connection and community that animate African ethical life by Fainos Mangena Read on Aeon·AI/AI 应用·aeon.co ↗aeon.co ↗
The benefits of medical AI assistance vary based on user expertiseStudy finds non-experts deferred to LLM-based diagnostic assistance, even when it was wrong, while clinicians caught AI errors.·AI/AI 应用, LLM·news.mit.edu ↗news.mit.edu ↗
阿里云容器服务Agent开启商业化收费36氪获悉,阿里云宣布,容器服务Agent将于北京时间9月3日10时增加新的技能中心和IM频道等功能并开启商业化收费。本次调整后,相关对话、自主智能运维及定时任务等功能将开始收费。调整过程不会影响存量用户已有设置和历史任务记录。 容器服务Agent以积分(Credit)作为计量单位,免费开通,支持按量付费及预付费资源包抵扣。使用容器服务Agent的对话、自主智能运维及定时任务等功能时产生的模型资源消耗会折算为积分进行计费。·AI/Agent, 智能体·36kr.com ↗36kr.com ↗
腾讯 WorkBuddy / CodeBuddy:Hy3 模型限时免费活动再次延长至 8 月 31 日IT之家 8 月 4 日消息,腾讯旗下的 CodeBuddy 和 WorkBuddy 今日宣布,CodeBuddy 和 WorkBuddy 中 Hy3 模型调用限时免费活动 延长至 2026 年 8 月 31 日 。 据IT之家此前报道,今年 7 月, 腾讯就将该限免活动延长过一次 ,当时延长至 8 月 5 日。 腾讯混元 Hy3 模型于 7 月 6 日发布并开源 。Hy3 是一个快慢思考融合的模型,采用 MoE 架构,总参数 295B·AI/大模型, AI 应用·ithome.com ↗ithome.com ↗
Kimi K3与DeepSeek V4之间,隔着原生多模态的时间差文 | 李炤锋 编辑 | 张雨忻 “长链任务如果只通过代码层面的反馈,误差可能会不断累积,最终效果会非常差。”谈及原生多模态的意义,一位多模态研究员表示,“视觉是一种更准确的反馈,也更贴近用户意图。” 过去一年,Coding与Agent能力不断改写大模型的排名,也成为AI最快兑现商业价值的场景之一。与此同时,随着Agent开始接管更多长链任务,越来越多的通用大模型开始匹配原生多模态能力。 今年1月,OpenAI发布的一份企业使用报告显示·AI/AI 应用, Agent, 多模态, +2·36kr.com ↗36kr.com ↗
New ways to learn and teach with ChatGPT Work and CodexExplore new education plugins for ChatGPT Work and Codex that help K–12 teachers, college educators, and students learn, teach, research, and build.·AI/AI 应用·openai.com ↗openai.com ↗