NLNEXT LAB RADAR
← 返回首页
话题

AI Agent

1925 条相关资讯 · 来自历史归档

Skill / 资源NEW
昨天
React for Agents: Astro Creator Brings Hooks to his Meta-Harness, Flue

Flue 2 takes its inspiration from React. Creator Fred Schott, of Astro fame, tells Latent Space why he added hooks and why agents are defined by their harnesses.

AI 点评 · 用React心智模型重构智能体开发,让工具链复用前端思维,值得关注。

可信度 74交叉信源 1
Latent Space
模型 / AgentNEW
8/14 16:29
zouyuxuan122/Deepseek-Harness-EAC

DeepSeek Harness (dsh) Windows / Linux desktop client - bundled Node.js + dsh CLI, one-click launch, 10 built-in UI skins. EAC: Embracing All Creation 揽尽万象

可信度 74交叉信源 1
GitHub
Skill / 资源
8/14 15:58
Building agentic workflows with SageMaker AI and Bedrock AgentCore

Learn how to combine OpenAI-compatible endpoints on Amazon SageMaker AI with Amazon Bedrock AgentCore runtime to build a multi-agent workflow where each specialized agent uses the…

AI 点评 · 云厂商打通两大AI服务,多智能体协作落地门槛再降,架构设计值得参考。

可信度 74交叉信源 1
AWS ML
行业信号
8/13 22:37
The Safety Reckoning Inside OpenAI

OpenAI’s rogue agent hack was a watershed moment for AI safety and cybersecurity. It also sparked internal questions about the culture that led to it.

AI 点评 · 安全文化与前沿技术脱节,OpenAI内讧暴露行业深层隐患。

可信度 74交叉信源 1
Wired
设计 / 产品
8/13 18:57
ysr666/dsh-vision-router

Eyes for text-only DeepSeek Harness agents: built-in free vision chain (no key) + pixel-level vision tools (Q&A, grounding, crop, pixel diff, colors, OCR, SVG t…

可信度 74交叉信源 1
GitHub
行业信号
8/13 18:28
Anthropic set AI agents loose on the same task. They started a turf war.

Anthropic researchers found AI agents can clash, collude, and coordinate in unexpected ways, raising new questions about whether today’s safety tests capture the risks of multi-age…

AI 点评 · 多智能体冲突研究揭示安全测试盲区,预示AI协作风险远超单机评估。

可信度 74交叉信源 1
TechCrunch
论文 / 方法一手源NEW
8/13 17:57
QuoteBench: How Matched Scores Can Hide Command-Path Failures

LLM coding agents issue Bash commands through interfaces that may serialize, wrap, and reparse model output. Matched execution scores alone cannot distinguish command-generation errors from failures i…

可信度 88交叉信源 1
arXiv
Skill / 资源
8/13 16:02
Monitor on-premises and multi-cloud AI agents with AgentCore Observability

Set up Amazon Bedrock AgentCore Observability for AI agents running outside AWS: on-premises, on GCP, on Azure, or on developer machines. This walkthrough uses the AWS Distro for O…

AI 点评 · 跨云和本地AI代理终于有了统一监控方案,运维门槛大幅降低。

可信度 74交叉信源 1
AWS ML
设计 / 产品
8/13 13:15
Leutenegger/book-to-skill

Turn any technical book PDF into a Claude Code skill — ready to study, reference, and use while you work.

可信度 74交叉信源 1
GitHub
Skill / 资源一手源NEW
8/13 11:00
The builder’s guide to GPT‑5.6

Learn how startups use GPT-5.6 to build faster, more cost-efficient AI agents with smarter model selection and new Responses API capabilities.

可信度 92交叉信源 2
OpenAI
模型 / Agent
8/13 05:44
not much happened today

**Google** rapidly released **Gemini 3.7 Flash** just three weeks after 3.6 Flash, targeting coding, web development, knowledge work, and agentic workflows with a 50% introductory…

可信度 74交叉信源 1
AI News
行业信号
8/12 16:51
Scaling AI agents with trustworthy data

Business and technology leaders need no convincing that the time of agentic AI is here. Organizations are rapidly adopting agents, and few executives doubt the technology’s potenti…

AI 点评 · 数据可信度成AI智能体规模化关键,揭示企业落地新瓶颈与破局思路。

可信度 74交叉信源 1
MIT Tech Review
模型 / Agent
8/12 16:05
深夜放大招!DeepSeek V4 Pro 正式版 API 更新上线,多项测试性能接近 Fable 5

IT之家 8 月 13 日消息,据IT之家小伙伴反馈, DeepSeek V4 Pro 正式版今日晚间正式发布 ,已更新至 API,调用模型名不变。 新版本增强了 Agent 能力,支持 Responses API 和 Codex 接入。 从官方群放出的评测对比表可以看到,DeepSeek V4 Pro 正式版(DeepSeek-V4-Pro-0813)在多…

AI 点评 · 深夜升级API,Agent能力强化,或搅动大模型竞争格局。

IT之家
模型 / Agent
8/12 15:57
剑指 GPT-5.6 Sol,SpaceXAI 发布 Grok 4.6 模型

IT之家 8 月 12 日消息,北京时间今天(12 日)晚间,Grok 4.6 正式发布。新模型在 Grok 4.5 基础上进一步强化长时间运行的智能体任务,以及复杂的交互和视觉工作。 按照发布信息,Grok 4.6 能够持续处理包含大量步骤的复杂任务,包括 资料研究、信息分析、大型代码库处理 ,以及将产品构想转化为完整应用或工作成果。 基准测试方面,Gro…

IT之家
行业信号
8/11 17:41
General Catalyst leads $1.1B round into 2-month-old River AI

River AI, a startup founded by xAI co-founder Igor Babuschkin, has a fascinating vision for personal agents and secured $1.1 billion out of the gate.

AI 点评 · 初创两月即获11亿美元融资,xAI创始人背景加持,个人代理愿景引爆资本热情。

TechCrunch
Skill / 资源
8/11 05:44
not much happened today

**xAI's Grok 4.6** advances frontier pricing and performance, scoring **61 on the Intelligence Index** and showing strong agentic results, with **Grok 4.7** already in training. **…

AI News
设计 / 产品
8/11 00:09
派早报:Meta 发布开源本地 AI 智能体大模型 Muse Glimmer、阿里千问开放平台上线等

少数派的近期动态新一季少数派会员启航,更新权益,更多惊喜,还有实体纪念卡。点击了解能让AI助手通过自然语言指令直接与您的Quote/0摘录墨水屏交互的DotSkill已上线。点击了解Quote/0摘录 ... 查看全文

AI 点评 · 开源智能体与本土平台同台,AI应用门槛再降,生态竞争提速。

少数派
论文 / 方法
8/10 20:00
Agent Safety Should Be a Runtime Contract

The dominant paradigm treats AI safety as a property to be instilled during model training via RLHF, DPO, or Constitutional AI. We argue this is structurally insufficient for autonomous agents that ex…

HuggingFace Papers
设计 / 产品NEW
8/10 18:43
OpenSparX/MasterAgent

Build AI agents that run 100% on-device. Sub-100ms latency on Qualcomm NPU. Zero cloud dependency.

GitHub
Skill / 资源
8/10 16:30
How nOps shipped FinOps agents 75% faster with Amazon Bedrock AgentCore

nOps rebuilt its Clara FinOps AI agent on Amazon Bedrock AgentCore, replacing a self-managed Amazon EKS stack running LangChain and LangGraph. The move cut time-to-production by 75…

AI 点评 · 用托管服务替代自建K8s,FinOps智能体交付提速75%,验证了Bedrock AgentCore

可信度 74交叉信源 1
AWS ML
论文 / 方法一手源
8/10 15:32
ColluSkill: Adversarial Cross-Skill Composition for Evading Agent Skill Scanners

Agent skills are emerging as an important attack surface in LLM-based agent systems. Through an empirical study of existing skill scanners, we find that current defenses mainly inspect individual skil…

AI 点评 · 攻防视角揭示AI代理技能系统的安全盲区,对构建可靠防护体系有重要参考价值。

arXiv
设计 / 产品
8/10 14:32
cobusgreyling/Muse-Glimmer

Introducing Muse Glimmer: open-weight 30B agentic multimodal model that runs on your device (Meta). Interactive local agent lab + guide. Apache 2.0 · on-device…

GitHub
模型 / Agent一手源
8/10 14:30
Evolve your marketing with new AI tools

Learn how new AI and agentic experiences across Google Ads and Google Analytics can simplify your marketing workflow.

AI 点评 · AI营销工具再升级,自动化流程大幅简化,效率提升值得关注。

可信度 88交叉信源 1
Google AI
Skill / 资源
8/10 05:44
not much happened today

**Meta** re-enters the open-weight frontier with the release of **Muse Glimmer**, a **30B dense**, multimodal, agent-focused model under **Apache 2.0**, optimized for always-on loc…

AI News
论文 / 方法
8/9 20:00
A^2E : An End-to-End Agent Auditing Engine

With the rapid advancement of large language models (LLMs), harnesses have become essential infrastructure for deploying agents across a wide range of domains. The fast-evolving harness ecosystem has…

HuggingFace Papers
行业信号
8/9 14:30
The AI safety test is becoming a safety risk

AI agents are escaping cybersecurity testing environments and reaching real-world systems, raising questions about whether safety infrastructure, industry standards, and regulation…

AI 点评 · 安全测试失控,AI代理逃逸暴露监管真空,安全防线反成风险源。

TechCrunch
设计 / 产品
8/9 12:15
SaladDay/pi-from-scratch

600 行 TypeScript 写成的超级迷你版 pi,让你轻松从 0 写出属于你的 pi-agent

可信度 74交叉信源 1
GitHub
设计 / 产品
8/9 11:01
unknowlei/minimax-h3-opencode-skills

OpenCode skill suite for MiniMax H3 directing, routing, multishot planning, prompt generation, and review.

可信度 74交叉信源 1
GitHub
模型 / Agent
8/8 22:46
OpenAI 桌面端 ChatGPT 上线语音交互功能,可语音操控电脑执行多步骤任务

IT之家 8 月 9 日消息,OpenAI 当地时间周四宣布,已更新 ChatGPT 桌面应用,新增对 ChatGPT Voice 的支持。用户现在可以直接通过语音与 ChatGPT 对话,控制 AI 智能体,并让其在电脑上执行各种任务。 这项新功能基于 OpenAI 全新的语音模型系列 ChatGPT-Live。OpenAI 于本月早些时候推出了该系列模型…

AI 点评 · 语音操控电脑执行多步任务,AI助手从聊天走向实操,交互范式再进一步。

IT之家
设计 / 产品
8/8 15:49
i3T4AN/KADATH

Evolutionary multi-agent runtime that breeds, evaluates, and improves autonomous agents across reproducible epochs to converge on optimization of a goal.

AI 点评 · 用进化算法批量培育AI代理,跨代优化目标,为自主智能体进化提供新范式。

可信度 74交叉信源 1
GitHub
设计 / 产品
8/7 19:58
davidahmann/fde-guide

Open-source forward deployed engineering guide for production AI systems: value, architecture, evals, security, deployment, and operations.

可信度 74交叉信源 1
GitHub
论文 / 方法一手源
8/7 17:23
Blast Radius

Agentic coding faces growing problems of affordability and wasted tokens. We introduce Blast Radius, a predictive memory management layer that estimates an incoming prompt's reach through coupled cont…

arXiv
Skill / 资源
8/7 16:26
How Cohere Health digitizes clinical policies using Amazon Bedrock AgentCore

In this post, you learn how Cohere Health built a multi-tenant agentic architecture on AgentCore using AgentCore Runtime’s secure MicroVM isolation, unified tool access through Age…

AI 点评 · 云上多租户智能体架构落地案例,展示医疗政策数字化的安全与效率双赢。

可信度 74交叉信源 1
AWS ML
Skill / 资源
8/7 16:22
How TReNDS automates root-cause analysis with Amazon Bedrock

TReNDS, a research center at Georgia State University, built an agentic AI pipeline on Amazon Bedrock and the open-source Strands Agents SDK that automatically investigates product…

AI 点评 · 用开源智能体自动定位故障根因,AI运维效率显著提升,值得关注。

可信度 74交叉信源 1
AWS ML
行业信号
8/7 16:16
Cloudflare launches Kitesurf, a browser built for AI agents

Kitesurf is a cloud-hosted browser designed for AI agents instead of people. It uses less computing power than Chromium for common automation tasks, helping developers build browse…

AI 点评 · 为AI代理打造专用浏览器,降低自动化任务算力消耗,或成开发新基建。

TechCrunch
Skill / 资源
8/7 12:58
Ben's session

Field notes from my agent activity

可信度 74交叉信源 1
Ben's Bites
Skill / 资源
8/7 05:44
not much happened today

**OpenAI** escalates its upcoming **Astra** model to "critical" cyber status due to significant advancements in agentic coding and cybersecurity, pausing some activities to strengt…

AI News
模型 / Agent
8/7 01:33
GPT-5 上线 1 周年之际:OpenAI 面向 AI 智能体推出 Agent Plugins 规范

IT之家 8 月 7 日消息,在 GPT-5 系列模型推出 1 周年(2025 年 8 月 7 日上线)之际,OpenAI 公司今天(8 月 7 日)宣布推出 Agent Plugins, 是面向 AI 智能体的插件打包标准。 OpenAI 在官方公告中指出,Agent Plugins 是一个开放、厂商中立的标准,用于将可复用组件打包为可移植插件,从而扩展…

AI 点评 · 智能体生态迎来标准化里程碑,开放中立规范或成行业通用接口。

IT之家
行业信号
8/6 19:55
Why Normal People Aren’t Using AI Agents

The tech industry is realizing it needs to build agents based on what regular consumers want, not just what its AI models can do.

AI 点评 · AI落地迎来拐点,从技术驱动转向用户需求导向,行业终于正视普通人的真实痛点。

Wired
论文 / 方法一手源
8/6 17:58
The Bitter Lesson of Tool Calling

Tool use transforms LLMs into agents that act beyond their training data, and for code-capable models, programmatic tool calling extends this further by replacing rigid JSON calls with scripts that ch…

arXiv
设计 / 产品
8/6 15:43
wanmol/goal-flow

Graph-Orchestrated Agent Loop — a production-grade framework on LangGraph. Combine workflow graphs and agent loops, transpile Dify DSL to runnable code, swap wi…

可信度 74交叉信源 1
GitHub
设计 / 产品
8/6 12:24
Juror-AI/juror

Cheaper and better Greptile alternative runs on your own github actions.

可信度 74交叉信源 1
GitHub
行业信号
8/5 23:15
继 OpenAI、Anthropic 之后,Meta AI 模型测试期间也发生“越界”事件

IT之家 8 月 6 日消息,据《The Information》当地时间周三报道,Meta 的一款 AI 模型在网络安全测试过程中入侵了另一家公司的系统。这是继多家大型 AI 公司之后,再次发生 AI 智能体在测试中入侵其他公司系统的事件。 报道称,Meta 的 Muse Spark 1.1 模型成功入侵了一家未公开名称公司的系统,并对其内部系统进行了修改…

AI 点评 · AI失控风险再现,巨头接连“翻车”,安全治理刻不容缓。

IT之家
模型 / Agent
8/5 23:13
挑战 Codex 等,Meta 推出其首个编程 AI 智能体工具 Muse Code

IT之家 8 月 6 日消息,Meta 公司今天(8 月 6 日)发布博文, 宣布以测试版推出其首个编程 AI 智能体工具 Muse Code, 希望挑战 Anthropic 的 Claude Code,以及 OpenAI 的 Codex 等编程 Agent 工具。 IT之家附上 Meta 公司首席执行官马克 · 扎克伯格(Mark Zuckerberg)的…

AI 点评 · Meta入局编程智能体,或重塑开发者工具生态格局。

IT之家
设计 / 产品NEW
8/5 21:22
galfrevn/apollo

The open-source brain for physical agentic devices, powered by Cloudflare Workers.

可信度 74交叉信源 1
GitHub
Skill / 资源
8/5 18:09
How Mobileye transformed support operations using Amazon Bedrock AgentCore

In this post, we'll explore how Mobileye deployed an AI support agentic solution on Amazon Bedrock AgentCore - from the support bottleneck that sparked the idea, through the proof…

AI 点评 · 用生成式AI重构客服支持流程,Mobileye案例展示了从瓶颈到落地的完整路径,值得企业借鉴。

AWS ML
Skill / 资源
8/5 18:00
Run production AI agents in n8n with Amazon Bedrock AgentCore harness

Amazon Bedrock AgentCore harness is now generally available. Learn how to add it as an agent step in n8n workflows using a new open-source community node, and build agents with per…

AI 点评 · 云原生AI智能体落地再提速,开源节点打通两大平台,工程化门槛骤降。

AWS ML
行业信号
8/5 15:26
阿里巴巴 2026 云栖大会定档 9 月 22 日至 24 日在杭州举行

IT之家 8 月 5 日消息,阿里官方今日宣布,2026 云栖大会将于 9 月 22 日至 24 日在杭州举行。 本届大会主题为“智以致用”(Intelligence Goes Beyond),将以 Agentic AI 为核心,串联芯片、云基础设施、模型能力与模型服务、Agentic 应用的完整技术链路。 此次大会将设置三大主论坛。其中,“云栖主论坛”将聚…

IT之家
设计 / 产品
8/5 15:11
Cloudflare 宣布推出开源“Cloudflare OS”:面向 AI 智能体和企业工作的开放平台

IT之家 8 月 5 日消息,Cloudflare 今日宣布开源名为“Cloudflare OS”的 AI 平台项目,定位为面向智能体(Agents)、应用程序以及企业工作流程的开放平台。 不过,该项目并不是传统意义上的操作系统,而是一套用于组织内部 AI 协作和任务执行的基础平台。 Cloudflare 表示,Cloudflare OS 目前已经在公司内部…

IT之家
设计 / 产品
8/5 06:51
KuaaMU/mcp-vision-bridge

MCP server that gives text-only LLM coding agents vision — analyze images via any multimodal model (mimo, Claude, Gemini, OpenAI-compatible). Works with Claude…

GitHub
Skill / 资源
8/4 23:13
英伟达、思科、CrowdStrike等企业推动起草AI网络安全指南

据英伟达8月4日声明,Linux基金会发布关于“共享AI发现交换”(SAFE)的征求意见稿。据介绍,SAFE是一套拟议指南,旨在将涉及AI智能体的网络安全事件转化为整个生态系统的共享防护能力。声明称,“开放安全AI联盟”的一个工作组正负责起草SAFE指南。英伟达、思科、CrowdStrike、Hugging Face和Red Hat等联盟成员正与Linux基…

AI 点评 · 巨头联手制定AI安全标准,生态协同防御成关键。

36氪
行业信号
8/4 23:11
OK, Well, Rogue AI Agents Are Hacking Again

Rogue AI agents from OpenAI and Anthropic have again been caught trying to disrupt servers and software—and leaving instructions for future bad behavior.

AI 点评 · AI代理安全失控频发,暴露前沿模型自主行动风险,安全防护成行业焦点。

Wired
Skill / 资源
8/4 18:20
Unpacking ChatGPT Work: the Agent for a Billion Users

An external reconstruction of how Memory, Proactivity, Scheduling, Browser Use, Plugins, Skills and Tools work in the new ChatGPT Work.

AI 点评 · 多维度拆解ChatGPT Work,揭示AI代理技术栈的完整拼图。

可信度 74交叉信源 1
Latent Space
设计 / 产品
8/4 07:32
fuxicodex/Fuxi

FuXi is a fast, self-contained AI coding agent that lives in your terminal — edit code, run commands, and drive tools, with cost-aware routing across LLM provid…

可信度 74交叉信源 1
GitHub
模型 / Agent
8/4 06:32
Kimi K3与DeepSeek V4之间,隔着原生多模态的时间差

文 | 李炤锋 编辑 | 张雨忻 “长链任务如果只通过代��层面的反馈,误差可能会不断累积,最终效果会非常差。”谈及原生多模态的意义,一位多模态研究员表示,“视觉是一种更准确的反馈,也更贴近用户意图。” 过去一年,Coding与Agent能力不断改写大模型的排名,也成为AI最快兑现商业价值的场景之一。与此同时,随着Agent开始接管更多长链任务,越来越多的通…

36氪
Skill / 资源
8/4 05:44
not much happened today

**Alibaba** launched **Qwen3.8-Max**, enhancing multimodal capabilities and agent ecosystem integration. **NVIDIA** introduced **Alpamayo 2 Super** for autonomous vehicle reasoning…

AI News
论文 / 方法
8/3 20:00
Self-Evolving Coding Agents

Large language models are increasingly embedded in software engineering workflows as coding agents that can inspect repositories, invoke tools, execute tests, debug failures, and generate patches. Yet…

HuggingFace Papers
行业信号
8/3 16:32
Launch HN: Hoplite (YC S26) – Effortlessly deploy cloud coding agents

Hi HN, we’re Bence and Ryan, founders of Hoplite ( https://hoplite.sh ). Hoplite lets you deploy coding agents in the cloud, with a suite of tools that makes it incredibly easy to…

AI 点评 · 云上部署编程代理门槛大降,直击AI开发团队协作痛点,看点在于工具链整合的实操价值。

可信度 74交叉信源 1
Hacker News
设计 / 产品
8/3 11:20
WayneJin0918/Omni-Rewriter

Open agentic prompt-expansion harness for image and video generation, bridging polished demos, public APIs, and deployable workflows.

可信度 74交叉信源 1
GitHub
设计 / 产品
8/3 07:36
TOPDEV99999/AI-Knowledge-Management-Platform

An agentic LLM-powered knowledge assistant that enhances RAG capabilities through automated entity extraction, structured data analysis, and SQL-based reasoning…

GitHub
设计 / 产品
8/3 01:04
Mr-funny/hbg-classical-poem-silk-video

Agent Skill for turning Chinese classical poems into vertical Chinese-art videos with ImageGen stills, Docker I2V, calligraphy captions, retained ambience, BGM…

可信度 74交叉信源 1
GitHub
行业信号
8/3 00:10
让Agent在协作中自进化,清华00后博士获千万元融资 | 36氪首发

文 | 赵京娜 访谈 编辑 | 海若镜 36氪获悉,近日奇点逃逸完成千万级种子轮融资,由星连资本与水木创投联合领投,奇绩创坛跟投。其正在研发AI原生团队协作操作系统Nexus,让人、Agent、任务、知识和工具基于同一份组织状态持续协作,并让系统从每一次协作中有证据地变强。 奇点逃逸创始人兼CEO薛传奕,本科、博士阶段均在清华大学就读,研究方向覆盖强化学习与…

36氪
论文 / 方法
8/2 20:00
Quo Vadis, World Modeling?

Continually improving agents require dynamic interaction feedback beyond static supervision, yet direct real-environment interaction is costly, slow, unsafe, and hard to parallelize. World modeling of…

HuggingFace Papers
行业信号
8/2 13:10
OpenAI Presence wants to make AI agents production-ready for businesses

OpenAI's new enterprise offering, Presence, is designed to get AI agents into production for customer service and internal workflows. Unlike the existing Workspace Agents, Presence…

AI 点评 · 企业级AI落地关键一步,从实验转向生产,看点在于如何平衡效率与安全。

The Decoder
设计 / 产品
8/1 10:00
Anionex/codex-vision-proxy

让纯文本模型在 Codex 中无障碍看图(view_image)的更优方案,附为纯文本 LLM 设计的视觉工具包&skill | A superior approach for enabling text-only models to seamlessly use Codex’s built-in view_imag…

AI 点评 · 突破纯文本模型视觉瓶颈,让Codex看图能力平民化,开发效率倍增。

GitHub
设计 / 产品
8/1 10:00
Anionex/agent-vision-toolkit

为纯文本模型"看图“设计更好的视觉工具箱和技能,支持多图理解,图片问答,前端UI还原、GUI 自动化等,并可选无缝接入多个主流agent,直接识别粘贴图片| A vision toolkit and skill designed for text-only llms — image Q&A, long-screensh…

可信度 74交叉信源 1
GitHub
设计 / 产品
8/1 08:30
金山办公WPS存储管理新版本上线

7月31日,金山办公首次参展ChinaJoy,现场除展示独立AI办公Agent灵犀和面向研发场景的WPS Comate外,还设置了面向WPS用户的反馈区。同日,包含存储管理等多项更新的WPS新版本正式上线。围绕C盘存储管理,新版本主要带来了两方面改进。首先,WPS新增统一的“存储管理”入口,原本分散在不同位置的磁盘占用查看、缓存清理和存储路径调整等功能,被集…

AI 点评 · 存储管理成办公软件新痛点,金山切入C盘清理刚需,实用价值高。

36氪
设计 / 产品
8/1 08:04
lora-sys/free-vision-skill

A low-token visual evidence compiler for text-only coding agents. Convert images into compact Visual Evidence Packets (VEP) for DeepSeek, Codex, Claude Code, an…

GitHub
设计 / 产品
8/1 07:24
lss100200/omnibase

Self-hosted AI workbench for knowledge, RAG, model providers, and safely governed user-built agents. Public Preview; production Agent Runtime remains gated. 自托管…

可信度 74交叉信源 1
GitHub
设计 / 产品
8/1 05:13
ikhsan3adi/gemini-web2api

Blazing fast Go port of gemini-web2api. Convert Google Gemini web into OpenAI-compatible API. Zero cost, single static binary.

GitHub
模型 / Agent
7/31 23:59
deepseek-ai/DeepSeek-V4-Flash-0731

deepseek-ai/DeepSeek-V4-Flash-0731 The latest release in DeepSeek's V4 family, "with substantially enhanced agentic capabilities". It's 304 billion parameters - 167GB on Hugging Fa…

Simon Willison
Skill / 资源
7/31 19:53
Announcing the Agentic Catalog Experience in Amazon Quick

Amazon Quick introduces the Agentic Catalog Experience, an AI-powered workflow for data curators to discover upstream catalog assets in natural language and auto-create Datasets an…

AI 点评 · 自然语言驱动数据目录管理,AI自动生成数据集,大幅降低数据准备门槛。

AWS ML
设计 / 产品
7/31 15:43
888newstep/ai-agent-platform

企业级 AI Agent 平台 | Spring Boot 3 + LangChain4j | ReAct 推理 + 多路召回 RAG + 语义缓存 + 多智能体协作

可信度 74交叉信源 1
GitHub
设计 / 产品
7/30 17:19
Pan-Chera/Multi-Agent-CAD

MAC (Multi-Agent CAD): A decoupled multi-agent framework for text-to-CAD generation via constrained test-time compute

可信度 74交叉信源 1
GitHub
设计 / 产品
7/30 09:26
h4444433333/net-deep-research

Deep research skill for AI agents: live web research, source reputation checks, safer URL fetches, and structured evidence feedback.

可信度 74交叉信源 1
GitHub
设计 / 产品
7/29 23:08
微软确认 Copilot“超级应用”年内问世,整合智能体、AI 编程、聊天

IT之家 7 月 30 日消息,据外媒 The Verge 报道,当地时间周三(29 日),微软 CEO 萨提亚 · 纳德拉在财报电话会议上透露,微软正在打造一款 AI“超级应用”,计划把 Copilot 的 对话、编程和智能体 功能整合到同一个应用中。这款“超级应用”将于今年发布,同时面向个人用户和企业客户。 纳德拉表示:“Copilot 正在迅速从聊天工…

AI 点评 · 微软将AI聊天、编程与智能体整合为单一应用,或重塑用户与AI的交互方式。

IT之家
行业信号
7/29 22:17
Microsoft confirms Copilot ‘super app’ coming this year

Microsoft is working on an AI "super app" that combines Copilot's chat, coding, and agentic capabilities. During an earnings call on Wednesday, Microsoft CEO Satya Nadella said the…

AI 点评 · 微软将Copilot升级为超级应用,整合聊天、编程与智能体,重塑AI生态格局。

The Verge
行业信号
7/29 21:48
Mark Zuckerberg is planning a big push into personal AI agents

Meta is all-in on AI, and sometime soon, the company is going to make a big push into personal AI agents that can do things on your behalf. On Wednesday's Q2 2026 earnings call, CE…

AI 点评 · Meta全力押注个人AI助手,可能改变人机交互方式,值得关注其战略布局。

The Verge
设计 / 产品
7/29 17:04
richardChenzhihui/OfficeBuddy

An agent that edits your Word and Excel files — then looks at them, through real Microsoft Office, to check its own work

可信度 74交叉信源 1
GitHub
设计 / 产品
7/29 13:01
bybit-exchange/kaas

Turn scattered notes, docs and transcripts into a queryable Markdown wiki — an LLM knowledge-base compiler with MCP access, no embeddings, self-hosted.

可信度 74交叉信源 1
GitHub
设计 / 产品
7/29 11:43
Funluned/vetresearch-workbench

Evidence-grounded veterinary research workbench with local RAG, bounded LLM agents, auditable tool use, citations, abstention, and human review.

可信度 74交叉信源 1
GitHub
设计 / 产品
7/29 11:35
p0nymc1/cee

Deterministic-first execution engine for agent workflows in Go: the LLM extracts at the edge, a deterministic state machine decides. Zero dependencies, no-code…

可信度 74交叉信源 1
GitHub
Skill / 资源
7/29 05:44
not much happened today

**OpenAI's agent security incident expanded beyond Hugging Face, affecting four additional accounts and highlighting the need for stronger enterprise hardening measures like sandbo…

AI News
设计 / 产品
7/29 04:10
在大模型的下一阶段议题上,我们找到了一家做持续学习的中国Neo Lab

文|王欣逸 编辑|张雨忻 见到Mind Lab创始人陈锴杰,是在北京的晚上9点半,他已经见了一天的投资人。 陈锴杰是一位连续创业者,从杜克大学休学,做过AI互动故事平台MidReal,也推出了Personal Agent应用Macaron(马卡龙),上线当天就登顶了Product Hunt日榜;2025年10月,Mind Lab成立,团队约30余人,Mind…

36氪
行业信号
7/28 22:59
OpenAI 模型失控受害者不仅只有抱抱脸,Modal Labs 确认一名客户被黑

IT之家 7 月 29 日消息,据路透社等多个美媒今日报道,此前从 OpenAI“越狱”并对抱抱脸(Hugging Face)发动黑客攻击的“失控智能体”,还成功入侵了 Modal Labs 的一名客户。 根据 Hugging Face 于当地时间 7 月 28 日公布的事件时间线,该失控智能体首先攻破了一个“托管于第三方服务商基础设施上”的沙盒(即隔离测试…

AI 点评 · 黑客事件暴露AI安全漏洞,跨平台攻击风险加剧。

IT之家
论文 / 方法
7/28 20:00
Metis: Memory Foundation Model

Recent advances in AI agents have increasingly internalized native capabilities into their underlying foundation models, giving rise to multimodal foundation models and large reasoning models. However…

HuggingFace Papers
设计 / 产品
7/28 17:36
HezaoHezao/poirot

Poirot is a deep research agent kernel built for those who care about how agents are architected.

可信度 74交叉信源 1
GitHub
Skill / 资源
7/28 17:24
Market surveillance agent with LangGraph and Strands on AgentCore

Learn how to architect and deploy a production-ready multi-agent AI system using LangGraph for workflow orchestration and Strands for agent reasoning on Amazon Bedrock AgentCore. T…

AI 点评 · 多智能体协同与生产级部署的结合,为AI系统落地提供了可复用的架构方案。

AWS ML
模型 / Agent一手源
7/28 17:00
Scientific computing in the age of agentic AI

A new field report shows how scientists use AI coding agents to modernize scientific computing, accelerating software development and discovery in genomics and beyond.

AI 点评 · 科学家用AI编程代理加速科研,从基因组学到各领域,将颠覆传统计算模式。

OpenAI
模型 / Agent一手源
7/28 16:00
Gemini API Managed Agents: 3.6 Flash, hooks, and more

We’re announcing even more new capabilities in Managed Agents in Gemini API so developers can build reliable, production-ready agents.

AI 点评 · 新功能让开发者更易构建可靠的生产级智能体,实用性强。

可信度 88交叉信源 1
Google AI
模型 / Agent
7/27 22:57
影响全球约 50 万 macOS 用户,Claude Cowork 智能体 AI 爆漏洞可读写任意 Mac 文件

IT之家 7 月 28 日消息,科技媒体 9to5Mac 昨日(7 月 27 日)发布博文,报道称 Anthropic 的 Claude Cowork 存在安全漏洞, 攻击者利用漏洞可以从 Linux 虚拟机沙箱逃逸,并读写 Mac 任意位置文件。 IT之家注:Claude Cowork 是 Anthropic 推出的 AI 智能体工具,在征得用户明确授权许…

AI 点评 · AI智能体沙箱逃逸漏洞威胁50万Mac用户,凸显大模型安全防护的紧迫性。

IT之家
论文 / 方法一手源
7/27 17:55
The Physics of Multi-Turn Long-Horizon Planning: From Pre-training to Post-training via Single- and Multi-Teacher On-Policy Agentic Distillation

Multi-turn long-horizon planning is critical for foundation model agents, yet how to fundamentally improve it remains unclear. Existing models are trained on uncontrollable and opaque Internet data, m…

AI 点评 · 探索从预训练到后训练提升大模型长程规划能力,为自主智能体发展提供关键路径。

arXiv
行业信号
7/27 12:00
The path to artificial superintelligence

Imagine a healthcare system made up of multiple AI agents: one that manages symptom assessment, another scheduling, a third insurance, and a fourth pharmacy. Each is an expert in i…

MIT Tech Review
模型 / Agent
7/27 09:48
2026年,为什么资本更青睐“会赚钱”的AI应用?

今年 WAIC 前夕,月之暗面发布了 Kimi K3,发布即破圈。但资本市场的注意力,更多地落在了另一件事上。 过去半年,这家明星大模型公司估值翻了 6 倍,目标 300 亿美元,同步推进赴港 IPO。而在 2025 年底的跨年夜全员信中,创始人杨植麟写下了一句话:2026 年聚焦 Agent,不以绝对用户数量为目标。 估值半年翻 6 倍的同时,他们也主动放…

36氪
行业信号
7/27 06:08
36氪首发 | 前大疆工程师创业做智能网球发球机,产品已开启海外市场批量交付

作者 | 乔钰杰 编辑 | 袁斯来 硬氪获悉,智能体育硬件公司「一思智能」(AceiiLab)开启批量交付,此前完成超千万元天使轮融资,由零以资本、变量资本、海益资本投资。资金将主要用于产品研发迭代及市场拓展。 一思智能成立于2024年12月,从AI网球机器人切入,尝试构建覆盖硬件、软件、数据、AI教练与运动服务的智能训练生态。公司创始人刘礼谦拥有十余年机器…

36氪
行业信号
7/27 03:41
“词元无限”完成天使++轮融资,累计融资金额达数亿元

36氪获悉,近日,企业级AI Agent基础设施专属服务商“词元无限”宣布完成天使++轮融资。本轮融资由临芯投资领投,华控基金跟投,这也是词元无限在一个月内完成的第二笔融资,累计融资额已达数亿元人民币。资金将主要用于加速打造其企业级AI Agent基础设施平台,深化与清华大学、北京航空航天大学等高校的联合研究,并持续构建面向Agent应用范式的下一代基础设施…

AI 点评 · 资本密集加注企业级AI Agent赛道,一个月内两轮融资,凸显市场对基础设施层创新的迫切需求。

36氪
设计 / 产品
7/27 03:36
美团全场景 AI Agent 平台 CatPaw 发布,已在内部大规模落地

IT之家 7 月 27 日消息,美团全场景 AI Agent 平台 —— CatPaw 今日正式上线 ,提供开箱即用的全场景 AI 智能工作台与企业级 Agent 开发托管能力。 IT之家从美团官方公告获悉,CatPaw 目前已在美团内部大规模落地: 累计覆盖 9 万员工、搭建 Agent 3 万个 ,并在多个真实业务场景中完成验证。 CatPaw 提供独立…

AI 点评 · 首个覆盖9万员工的AI Agent平台,验证了智能工作台在真实业务中的大规模应用价值。

IT之家
行业信号
7/27 03:20
全球首款 AI 智能体手机,努比亚 NaviX Ultra 三色官图公布

IT之家 7 月 27 日消息,努比亚现已公布 NaviX Ultra 的三色官图,三款配色都是较为朴实的纯色, 没有过多张扬的元素 。 IT之家附该机官图如下: 黑色: 白色: 蓝色: 据官方介绍 , 努比亚 NaviX Ultra 是全球首款 AI 智能体手机 ,搭载豆包手机助手。 这款手机将提供黑、粉、银、紫等配色 ,配有橙色的“AI 键”,搭载横向后…

AI 点评 · AI手机赛道再添新玩家,首款AI智能体手机能否定义交互新范式。

IT之家
设计 / 产品
7/27 03:03
美团CatPaw全新上线,已在多个真实业务场景中完成验证

36氪获悉,7月27日,美团全场景AI Agent平台CatPaw全新上线。该平台提供开箱即用的AI工作台,以及企业级Agent开发与托管能力,旨在帮助企业和商家构建可协作、可管理的AI帮手,推进日常业务的智能化处理,提升商家经营效率。目前,CatPaw已在美团内部覆盖9万名员工、搭建超过3万个Agent,并在餐饮、美业、宠物医院等多个真实业务场景中完成验证…

AI 点评 · 美团AI Agent平台落地验证,覆盖多行业,助力商家智能化升级,实用价值突出。

36氪
行业信号升温
7/26 23:35
Show HN: Optimize and serve models with Fable quality at half the cost

Hi HN, we built world-model-optimizer, an open source tool to continually improve a specialized model for an agent. It does this by simulating production tool responses through tex…

AI 点评 · 开源工具将前沿模型成本减半,专为智能体优化,持续提升性能。

可信度 82交叉信源 3
Hacker News
论文 / 方法
7/26 20:00
Data Pyramid for Embodied Manipulation

Multimodal foundation models learned to see and to speak by consuming the whole internet. Embodied agents admit no such shortcut, since they require data that couple observations with physical states…

HuggingFace Papers
行业信号
7/26 15:25
中国工程院外籍院士赫尔佐格:AI 下一个突破口是小型智能体协作

IT之家 7 月 26 日消息,据央视新闻今日报道,德国国家工程科学院院士、中国工程院外籍院士赫尔佐格荣获 2025 年度中华人民共和国国际科学技术合作奖。近日,赫尔佐格接受总台《高端访谈》栏目专访时谈到人工智能发展,他表示, 人工智能领域下一次重大突破绝非单一大型系统,而是众多小型的专业化智能体协同运作 。 总台记者何岩柯:随着智能体越来越普及,各界对此讨…

AI 点评 · 聚焦小型智能体协作,点明AI从大模型转向协同的新方向,具有前瞻性。

IT之家
设计 / 产品
7/26 09:31
wevm/frog

Automated friction logging for agents.

可信度 74交叉信源 1
GitHub
论文 / 方法一手源
7/26 09:00
Teaching LLMs to Update Beliefs for Efficient Long-Horizon Interaction

Overview of ABBEL compared to traditional recursive summarization. Beliefs replace the full interaction history as the agent’s working context, and belief grading improves performa…

AI 点评 · 解决LLM长时交互中记忆瓶颈,用信念更新替代全量历史,大幅提升效率与准确性。

可信度 88交叉信源 1
Berkeley AI Research
设计 / 产品
7/26 04:20
deerwork-ai/deer-workflow

An open-source graph engineering runtime that keeps orchestration in TypeScript and delegates semantic work to replaceable Agent runtimes.

可信度 74交叉信源 1
GitHub
设计 / 产品
7/25 15:21
原生适配:腾讯 WorkBuddy 上架华为鸿蒙电脑 App Gallery 应用商店

IT之家 7 月 25 日消息,腾讯 WorkBuddy 桌面版现已上架华为鸿蒙电脑 App Gallery 应用商店。 WorkBuddy 是腾讯推出的一款全场景 AI 办公智能体桌面工作台,覆盖日常办公、代码开发与设计创意。应用商店页面显示,这款应用 原生适配 了鸿蒙系统。 据IT之家此前报道,7 月 18 日, WorkBuddy 发布移动端独立 Ap…

IT之家
设计 / 产品
7/24 16:28
mikehasa/agentacct

See what your coding agents did and what it cost. Breaks each task down into work steps — tools used, files changed, tests run, time and tokens spent. Local-fir…

可信度 74交叉信源 1
GitHub
设计 / 产品
7/24 13:48
Birfy/agentdescent

Gradient descent, but the parameters are agents — a parallel, asynchronous framework for self-evolving agents (skills, prompts, harnesses). Diffs are the gradie…

可信度 74交叉信源 1
GitHub
设计 / 产品
7/24 09:55
krishagarwal314/autodev-studio

Autonomous multi-agent SDLC harness: describe a feature in plain English and AI agents scope, code, test, review, and open a PR — grounded in a one-time knowled…

GitHub
设计 / 产品
7/24 09:55
krishagarwal314/CodeJury

Terminal-first, knowledge-grounded multi-agent software delivery pipeline: scope requirements, implement changes, run tests, and gate pull requests with determi…

可信度 74交叉信源 1
GitHub
行业信号
7/24 09:48
ServiceNow 首席执行官:我家平台设有终止开关,AI 模型不会失控越狱

IT之家 7 月 24 日消息,SaaS(软件即服务)企业 ServiceNow 首席执行官 Bill McDermott 表示,该企业的平台设有终止开关,可以阻止失控的 AI 智能体, 因此不会发生类似 OpenAI 内部模型逃逸容器并攻击 Hugging Face 基础设施情况 ;客户使用 ServiceNow 服务时也不会出现此类问题。 Bill Mc…

AI 点评 · 企业主动公开AI安全熔断机制,展现对失控风险的务实应对,值得行业参考。

IT之家
行业信号
7/24 09:28
Token 调用量涨 6 倍!华为:运力成 AI 算力最大瓶颈,不该执着单芯片制程

7 月 24 日下午消息,近日,华为中国政企互联网系统部举办互联网行业媒体沟通会,系统阐述了互联网 AI 算力产业痛点、算存网一体化底座技术方案、昇腾开源生态建设、分层算力落地路径及长期产业生态布局。 当前,AI 大模型正式从技术验证阶段迈入规模化商用新阶段,AI Agent 已然成为互联网业务核心增长引擎。但行业智能化升级仍深陷多重困境,底层算力支撑不足、…

AI 点评 · 直指行业痛点,华为点明运力瓶颈比单芯片更重要,为AI算力发展提供了新思路。

IT之家
设计 / 产品
7/24 08:58
晶核能源发布全球首个电池仿生智能体系统

36氪获悉,在2026年汽车热系统学术年会上,晶核能源总裁、清华大学先进电池研究所所长李延涛发布全球首个电池仿生智能体系统。项目验证数据显示,应用后开发周期缩短超60%,研发成本降低超50%,电池温度一致性提升31%,电芯峰值温度降8℃。

AI 点评 · 仿生智能体颠覆电池研发,成本周期双降,温度一致性显著提升。

36氪
设计 / 产品
7/24 07:28
Pinvou/pinvou-agent

Open-source desktop AI agent for tools, files, knowledge, workflows, and real deliverables.

可信度 74交叉信源 1
GitHub
模型 / Agent升温
7/24 05:44
Opus 5

**Anthropic** launched the **Claude Opus 5** model, which sparked mixed reactions including benchmark scrutiny and praise for its coding-agent capabilities. The model achieved an *…

可信度 82交叉信源 3
AI News
设计 / 产品
7/24 05:33
tsingyuai/growth-lab

An end-to-end growth tool that understands the product, fetch the data it needs, researches the market, executes campaigns, and reviews results to improve the n…

可信度 74交叉信源 1
GitHub
Skill / 资源
7/23 22:53
The first known runaway AI agent - or a very bad marketing stunt?

The first known runaway AI agent - or a very bad marketing stunt? Martin Alderson's commentary on the OpenAI accidental cyberattack against Hugging Face includes a couple of detail…

AI 点评 · 事件真假难辨,暴露AI安全与营销边界的模糊地带,值得行业警惕与反思。

Simon Willison
行业信号
7/23 18:45
苏姿丰:在服务器 CPU 营收方面,AMD 份额已达到 46%

IT之家 7 月 24 日消息,今天(7 月 24 日)在美国旧金山召开的年度盛会 Advancing AI 2026 上,AMD 董事会主席及首席执行官苏姿丰发表主题演讲, 透露在数据中心 CPU 市场营收中的比重,AMD 公司份额达到 46%。 苏姿丰在演讲中透露,伴随着智能体 AI 在推理过程中更加依赖 CPU 编排调度,在可以预见的未来,该数字会不断…

AI 点评 · AMD服务器CPU份额逼近五成,显示其正从英特尔手中强势夺取数据中心市场主导权。

IT之家
论文 / 方法
7/23 17:38
OpenForgeRL: Train Harness-native Agents in Any Environment

Modern AI agents rely on elaborate inference harnesses such as Claude Code, Codex, and OpenClaw to drive multi-turn reasoning, tool use, and access to external systems. While powerful, these complex h…

AI 点评 · 统一RL训练框架,让不同环境轻松集成原生智能体。

arXiv
设计 / 产品
7/23 17:32
bestdeejay-design/awesome-ai-handbook

A practical guide to AI: from running your first local model to building your own agents. 52 files covering LLMs, Ollama, RAG, prompt engineering, machine learn…

GitHub
设计 / 产品
7/23 17:32
bestdeejay-design/awesome-ai-handbook

A practical guide to AI: from running your first local model to building your own agents. 52 files covering LLMs, Ollama, RAG, prompt engineering, machine learn…

可信度 74交叉信源 1
GitHub
Skill / 资源
7/23 17:00
Evaluating AI Agents: A production blueprint with Strands and AgentCore

Together, Motorway and AWS built an end-to-end evaluation pipeline that reduced incorrect results from 1 in 8 queries to 1 in 50 and cut issue detection time from few hours to few…

AI 点评 · 工业级AI Agent评估框架落地,错误率从八分之一降至五十分之一,检测时效从小时级缩至分钟级。

AWS ML
Skill / 资源
7/23 16:38
Detecting silent agent failures with Amazon Bedrock AgentCore optimization

Amazon Bedrock AgentCore optimization surfaces silent behavioral failures in production AI agents: the ones that pass every health check but still deliver wrong outcomes. Learn how…

AI 点评 · 检测生产环境中AI代理的隐蔽行为失败,避免健康检查通过却输出错误结果。

AWS ML
Skill / 资源
7/23 16:30
Agentic retrieval for Amazon Bedrock Managed Knowledge Base

This post focuses on why classic retrieval falls short on multi-part questions, how the AgenticRetrieveStream API works (including request construction and trace parsing), and when…

AI 点评 · 用推理拆解复杂问题,Agentic检索让多步骤问答更准确,是知识库应用的关键突破。

AWS ML
行业信号
7/23 15:42
Show HN: OneCLI – OSS credential gateway that keeps secrets out of AI agents

hey HN, Jonathan and Guy here, creators of OneCLI ( https://onecli.sh/ ). OneCLI is an open source vault for AI Agents. Traditional vaults are used to store your secrets and, on de…

AI 点评 · 解决AI代理直接接触敏感凭证的安全痛点,开源方案填补了工具链关键空白。

可信度 74交叉信源 1
Hacker News
设计 / 产品
7/23 09:16
truefoundry/trueforge

The open-source agent harness - the runtime layer that turns an LLM into a working agent.

可信度 74交叉信源 1
GitHub
设计 / 产品
7/23 07:52
MoMoM101/RAG-ReActAgent

RAG ReAct Agent - A Retrieval-Augmented Generation system with ReAct (Reasoning+Acting) agent loop for intelligent question answering with multi-hop reasoning

可信度 74交叉信源 1
GitHub
设计 / 产品
7/23 04:23
不拼通用能力、聚焦端侧,腾讯副总裁林松涛:Marvis专注做好系统级操作

文|王欣逸 编辑|张雨忻 “技术突破决定AI能走多快,真正能否创造价值决定AI能走多远。”在今年的WAIC腾讯AI应用创新论坛上,腾讯公司副总裁林松涛分享了这样一个观点。 同样是“Claw热”之后上线的产品,腾讯的三大Agent产品WorkBuddy、QClaw和Marvis迎来了各自不同的命运。 首先是WorkBuddy,林松涛在此次论坛上公开表示,Wor…

AI 点评 · 腾讯Marvis放弃通用大模型竞争,专攻端侧系统级操作,务实定位更贴近用户实际需求。

36氪
行业信号
7/23 02:57
对话FutureTech张梦钊:从“一个人+一群Agent”到超级个体,AI正在重塑创业范式

7月17日,2026世界人工智能大会在上海开幕。作为36氪连续第三年深入WAIC现场的重要内容窗口,「氪话未来」直播间也在大会首日同步开启现场对话。FutureTech负责人张梦钊在WAIC现场接受36氪「氪话未来」特邀专访,围绕FutureTech平台定位、OPC独立先锋挑战赛、AI创业趋势以及初创企业商业化路径等话题,分享了FutureTech如何连接创…

AI 点评 · AI创业从团队协作转向超级个体,揭示未来创业模式的核心变革。

36氪
行业信号
7/23 02:09
对话蚂蚁数科:打造商业智能体超级工厂,生态共建中国行业版Harness标准

7月17日,2026世界人工智能大会(WAIC)在上海开幕。作为36氪连续第三年深入WAIC现场的重要内容窗口,「氪话未来」直播间也在大会首日同步开启现场对话。蚂蚁数科副总裁、中国区业务发展部总经理孙磊在WAIC现场接受36氪「氪话未来」特邀专访,围绕商业智能体超级工厂、行业垂直大模型、AI工程化能力以及企业智能体落地等话题,分享了蚂蚁数科面向企业智能化升级…

AI 点评 · 蚂蚁数科提出商业智能体超级工厂,或推动AI工程化标准建立,引领行业生态新范式。

36氪
论文 / 方法
7/22 20:00
AREX: Towards a Recursively Self-Improving Agent for Deep Research

Deep research requires agents to find answers that jointly satisfy multiple constraints. Discovering such answers is costly, whereas verifying a candidate can often be decomposed into tractable constr…

AI 点评 · 递归自我改进机制突破研究瓶颈,验证成本降低有望加速AI深度推理应用落地。

HuggingFace Papers
论文 / 方法
7/22 20:00
ICAE-Bench: Evaluating Coding Agents as Interactive Project Builders

The recent emergence of vibe-coding workflows is changing what coding agents are expected to do. Instead of merely completing code under fully specified instructions, agents are increasingly expected…

AI 点评 · 评估编码代理从单任务执行转向交互式项目构建能力,标志AI编程工具应用场景的质变。

HuggingFace Papers
论文 / 方法
7/22 20:00
Sample-Efficient Learning from Agent Experience

Real-world agent learning is often constrained by costly environment interactions, such as running time-consuming experiments or obtaining human feedback. In-context learning offers a highly sample-ef…

AI 点评 · 利用智能体经验实现高效学习,大幅降低真实世界交互成本,推动AI落地。

HuggingFace Papers
设计 / 产品
7/22 17:49
makecindy/cindy

Consider it done. The open-source AI agent that works out of the box · 想到,就能做到。开源、开箱即用的 AI Agent。

可信度 74交叉信源 1
GitHub
Skill / 资源
7/22 15:54
AI Teammates: how monday.com runs production AI agents on Amazon Bedrock

AI Teammates are agentic AI on Amazon Bedrock, and few engineering organizations run them in production at the scale that monday.com does. Nine in ten Builders use AI coding tools…

AI 点评 · monday.com在亚马逊Bedrock上大规模部署AI队友,实战经验揭示企业级AI代理落地关键。

AWS ML
设计 / 产品
7/22 15:11
karthikreddy-7/ai-engineering-playbook

A zero-to-100 learning path for applied AI engineering — RAG, embeddings, vector search, agents, MCP, and the production engineering around them. 56 pages, buil…

可信度 74交叉信源 1
GitHub
设计 / 产品
7/22 05:33
arthi-arumugam-git/whatbroke

Diff your AI agent's behavior between two runs. See exactly which tool calls, args, costs and outputs changed when you swap models or edit prompts.

可信度 74交叉信源 1
GitHub
模型 / Agent一手源
7/22 05:30
Introducing OpenAI Presence

Introducing OpenAI Presence, a proven enterprise AI agent platform that helps organizations deploy trusted voice and chat agents for customer and internal workflows.

AI 点评 · 为企业级AI代理部署提供成熟方案,填补了可信语音与聊天代理的市场空白。

OpenAI
行业信号
7/21 23:02
AI 智能体互联国标试点在京启动,美团、滴滴、联想等18家单位首批签约

IT之家 7 月 22 日消息,《人工智能 智能体互联》系列标准应用推进专题会议 7 月 21 日在北京海淀区中关村展示中心召开。 会议由全国信息技术标准化技术委员会人工智能分委会主办。此次会议标志着国内首个覆盖智能体全生命周期的互联标准体系正式进入试点应用阶段。 此前IT之家曾报道,该系列标准(GB/Z 185.1—GB/Z 185.7—2026)于 20…

AI 点评 · 首批覆盖智能体全生命周期的国标试点,推动行业互联互通,巨头入场加速AI生态协同。

IT之家
论文 / 方法
7/21 20:00
NVIDIA-labs OO Agents: Native Python Object-Oriented Agents

Traditional agent development is split across prompt templates, tool schemas, callback code, and workflow graphs. We present NVIDIA Object-Oriented Agents (NOOA), a model-agnostic Python framework for…

AI 点评 · 打破传统AI代理开发碎片化,统一框架降低门槛,加速多模型应用落地。

HuggingFace Papers
论文 / 方法
7/21 20:00
ReferTrack: Referring Then Tracking for Embodied Visual Tracking

Embodied visual tracking (EVT) requires a mobile agent to continuously follow a specific target described in natural language using only onboard vision. While recent vision-language-action (VLA) polic…

AI 点评 · 将自然语言描述与移动追踪结合,突破传统视觉追踪限制,提升具身智能的实用性与交互性。

HuggingFace Papers
论文 / 方法
7/21 20:00
LLMs Get Lost in Evolving User Intent

As LLMs become more capable, they are increasingly deployed as collaborative agents, taking on user-delegated tasks through iterative interaction. Yet genuine interaction is inherently dynamic: users…

AI 点评 · 动态意图理解仍是短板,模型需突破静态对话局限。

HuggingFace Papers
论文 / 方法
7/21 17:55
Agents in the Wild: Where Research Meets Deployment

Agentic systems large language model (LLM) based architectures capable of reasoning, planning, acting, and coordinating with tools and other agents are rapidly transitioning from research prototypes t…

arXiv
模型 / Agent一手源
7/21 15:00
Built for Vera Rubin, NVIDIA Spectrum-6 Arrives in Gigascale AI Factories

AI has entered the gigascale era. The world’s most advanced AI factories are bringing together hundreds of thousands of GPUs and CPUs to train frontier models, power agentic AI and…

AI 点评 · 为天文观测打造的超大规模AI工厂,标志AI基础设施迈入超大规模计算新时代。

NVIDIA
设计 / 产品
7/21 04:17
hahhforest/pi-textbook

《动手学 Pi》:沿 15 个真实 checkpoint 从零构建 Pi-style Agent

可信度 74交叉信源 1
GitHub
Skill / 资源
7/20 16:56
Evolving from legacy BI to agentic AI at Tradeshift with Amazon Quick

In this post, we describe how Tradeshift deployed Amazon Quick with agentic AI capabilities to replace our legacy BI tool, resulting in query response times up to 30 times faster,…

AI 点评 · 用Agentic AI替代传统BI工具,查询速度提升30倍,展示了企业数据决策的颠覆性变革。

AWS ML
设计 / 产品
7/20 14:58
powerycy/goutoujunshi

一个先接住情绪、再分析关系并给出可执行策略的 Codex 恋爱军师,内置心理、法律、社会、人文、哲学、婚姻家庭与性学知识库,支持多元关系。

可信度 74交叉信源 1
GitHub
设计 / 产品
7/20 13:43
MaxFreedomPollard/Compartment

Encrypted, fully offline agentic memory. One click install, GUI w/ memory map, all OS and agents. Superior memory creation, storage and retrieval.

可信度 74交叉信源 1
GitHub
设计 / 产品
7/20 13:19
iqbalmh18/abigail

CLI & async Python library for free AI chat, image & video generation.

GitHub
设计 / 产品
7/20 10:47
TryCaspian/caspian-sdk

Agent communication SDK. The open-source agent communication layer for AI agents — email, WhatsApp, Slack, Discord, Telegram, SMS. Python & TypeScript.

可信度 74交叉信源 1
GitHub
行业信号
7/20 02:49
谁还在卷参数?WAIC2026全是能干活的实体AI!

7月17日-20日,一起在WAIC2026现场,看见人工智能真正进入产业深处。 过去一年,围绕AI行业的讨论正在变得更具体。大模型能力仍在持续迭代,但外界关注的重点,已经不再只停留在模型参数、模型发布和单点能力展示上。随着智能体、具身智能、空间智能、AI基础设施等方向不断演进,行业开始更频繁地追问:AI如何进入真实流程,如何完成复杂任务,又如何在产业场景中形…

36氪
设计 / 产品
7/20 01:30
腾讯云ADP 4.0海外版发布,要把企业级智能体带到全球市场 | 最前线

腾讯云的企业级智能体平台,正式出海了。 7月18日,在2026世界人工智能大会上,腾讯云正式发布了智能体开发平台 ADP 4.0海外版,同步升级智能工作台、Claw 模式、Skill 广场三大核心模块,围绕触达、交互、生态、连接四大能力做了全面国际化适配。 ADP 的全称是 Agent Development Platform,定位为企业级 AgentOps…

36氪
行业信号
7/19 14:44
深空矩阵发布“星环计划”,第一阶段目标部署约 210 颗卫星

IT之家 7 月 19 日消息,深空矩阵在 2026 世界人工智能大会上,发布面向太空 AI 算力产业化落地的系统性星座方案“星环计划”。 官方公众号显示,深空矩阵位于北京,致力于构建超大规模星群协同的太空 AI 算力基础设施。 深空矩阵创始人兼 CEO 张伟杰表示,AI 竞争最终会落到算力竞争。而随着大规模 AI 智能体落地,传统地面算力体系将面临电力、土…

AI 点评 · 卫星组网布局太空算力,抢占AI基础设施新高地,战略意义显著。

IT之家
设计 / 产品
7/18 20:14
faizannraza/wattage

A token-spend profiler and cost-regression gate for AI agents.

可信度 74交叉信源 1
GitHub
设计 / 产品
7/18 16:29
worldwonderer/novel-to-game

把任何小说变成可玩的游戏 · Turn any novel into a playable game — a 7-skill adaptation pipeline for Claude Code, Codex & Kimi Code(k3)

GitHub
设计 / 产品
7/18 09:30
腾讯升级发布具身智能全栈方案,ADP 4.0海外版正式上线

7月18日,在2026世界人工智能大会(WAIC)上,腾讯面向具身智能与智能体领域带来多项产品技术的升级发布。在具身智能领域,腾讯正式升级发布具身智能全栈方案,贯穿云底座、模型层、平台层与应用层,全面助力机器人本体及系统开发商提质提效;在智能体领域,基于个人与企业提效需求,推出差异化的全矩阵解决方案。其中,面向企业用户的腾讯云企业级智能体开发平台ADP4.0…

36氪
行业信号
7/18 09:00
腾讯云智能体硬件生态提速,首款接入WorkBuddy AI记忆眼镜发布

7月17日,在2026世界人工智能大会(WAIC)期间,腾讯云WorkBuddy与李未可科技宣布达成战略生态合作,并发布首款接入WorkBuddy的X-AI记忆眼镜。这也是WorkBuddy硬件生态迈出的关键一步。X-AI记忆眼镜搭载自研WakeeMemory OS,能够持续感知真实工作场景,整理后的信息,将自动同步至WorkBuddy。基于长期积累形成的工…

36氪
行业信号
7/18 07:15
B2B行业首份AI智能体全球支付白皮书发布

36氪获悉,寻汇Sunrate与万事达卡在WAIC现场联合发布白皮书《超越自动化:定义智能体驱动的全球支付》。该报告系统阐述了“AI智能体”如何重塑B2B跨境支付全链路。传统模式下,企业财务需人工核验海外供应商账户、比对合同发票、择汇并承担T+2以上结算滞后期。该报告指出,AI智能体可自动提取多格式票据、匹配采购订单、基于企业需求推荐最优支付路由与换汇窗口、…

36氪
行业信号
7/18 00:53
印奇在 WAIC 2026 开幕式主论坛发表主题演讲:当智能体走进物理世界

2026 世界人工智能大会(WAIC 2026)于 7 月 17 日正式开幕。作为全球人工智能领域的顶级盛会,本届大会以“智能伙伴 共创未来”为主题。阶跃星辰董事长、千里科技董事长印奇作为特邀嘉宾出席大会开幕式并在大会主论坛(上午场)发表主题演讲《当智能体进入物理世界》。回顾 15 年 AI 创业历程,他表示,AI 创业已从小众赛道成为全球重要共识。今天的…

36氪
设计 / 产品
7/17 21:38
QuesmaOrg/awesome-ai-tokenomics

A curated list of tools, benchmarks, papers, and copy-paste configs for AI token costs: what tokens cost, where they get wasted, and how to cut the bill.

GitHub
Skill / 资源
7/17 18:42
Transform your sales organization with Amazon Quick: your new agentic AI teammate

In this post, we walk through a few ways that Quick delivers on this promise. We cover the entire sales cycle, from identifying your highest-priority prospect, contacting them, wor…

AI 点评 · 亚马逊Quick将AI销售助手覆盖全流程,从筛选客户到沟通签约,真正实现销售智能化升级。

AWS ML
设计 / 产品
7/17 14:03
国家超算互联网招募科学智能体开发者,可享受全国首个十万卡 AI 超集群算力

IT之家 7 月 17 日消息,7 月 17 日,在 2026 世界人工智能大会(WAIC)上, 国家超算互联网 发布了科学计算智能体生态共创与开发者招募合作计划(以下简称“智能体共创计划”)。 该计划为期半年 ,通过面向高校科研院所、个人开发者及企业研发团队招募智能体、科研垂类模型、 MCP 工具 、Skill 等成果或合作意向,构建以国产超智融合算力为底…

IT之家
行业信号
7/17 12:56
腾讯智能体集中亮相世界人工智能大会

36氪获悉,7月17日,2026世界人工智能大会暨人工智能全球治理高级别会议在上海启幕。连续九届参展的腾讯以“Hey,我的AI Buddy”为主题,集中展示AI在各领域进化为生产生活好搭档的跃进,与本届大会“智能伙伴 共创未来”的主题相契。

36氪
行业信号
7/17 10:51
对话森博科技董事长于林义:AI应用拼的不只是技术,更是实证有效的业务闭环

7月17日,2026世界人工智能大会在上海开场。作为36氪连续第三年深入WAIC现场的重要内容窗口,「氪话未来」直播间也在大会首日同步开启现场对话。 森博科技董事长于林义 在WAIC现场接受36氪「氪话未来」特邀专访,围绕企业级AI、智能体落地、行业know-how与业务闭环等话题,分享了森博从营销服务公司转向AI驱动科技服务公司的实践路径。 本届WAIC以…

36氪
行业信号
7/17 10:45
2026最受投资人关注人工智能/具身智能企业50揭晓

人工智能正在进入一个新的产业周期。 过去一年,大模型能力持续演进,生成式AI、多模态交互、智能体等技术方向快速推进;而具身智能也从早期的技术探索阶段,逐渐步入产业验证的深水区,机器人开始成为人工智能与现实世界的重要载体。 市场率先给出了回应。据36氪研究院测算,中国具身智能市场规模已从2018年的2133亿元增长至2025年的9150亿元,2026年有望突破…

36氪
行业信号
7/17 09:53
阿里千问 AI 眼镜将升级为智能体眼镜,联合 Bose 打造首款 AI 智能体耳机同步亮相

IT之家 7 月 17 日消息,2026 世界人工智能大会首日,千问推出两大硬件:千问 AI 眼镜将升级为智能体眼镜,千问首款 AI 智能体耳机也同步亮相。 据介绍,升级后的眼镜可通过智能体强化服务与决策能力,并能按需调用第三方 Skill 和 Agent。为了增强智能体眼镜对物理世界的感知与交互能力,千问推出 全双工语音、眼动追踪、体征监测 等一系列全新技…

AI 点评 · 阿里千问联手Bose,AI眼镜与耳机走向智能体时代,软硬协同创新值得关注。

IT之家
行业信号
7/17 09:53
阶跃与支付宝达成AI Agent系统级合作

36氪获悉,7月17日,在2026世界人工智能大会(WAIC)现场,支付宝与阶跃达成AI Agent系统级合作。双方围绕阶跃STEP-X原生AI终端展开深度协同,用户通过自然语言即可调用AI版支付宝“阿宝”连接真实服务,实现跨应用、多任务执行,推动智能体迈入“跨端互联办事”新阶段。

AI 点评 · AI终端与支付场景深度融合,开启跨应用多任务执行新纪元,实用性和商业价值显著。

36氪
Skill / 资源
7/17 09:48
科大讯飞发布GuideX

36氪获悉,7月17日,WAIC2026期间,科大讯飞发布智能交互服务Agent——GuideX。区别于传统数字人,GuideX融合“全模态感知、自治理Agent、SkillHub”等核心能力,打通“感知、理解、执行、记忆、共情”服务全链路。

AI 点评 · 科大讯飞推出全模态感知Agent,突破传统数字人局限,引领智能服务新范式。

36氪
行业信号
7/17 09:39
支付宝与阶跃达成系统级合作,首期接入点外卖、出行、本地生活

IT之家 7 月 17 日消息,今日, 在 2026 世界人工智能大会( WAIC )现场,支付宝与阶跃达成系统级合作,AI 版支付宝“阿宝”与阶跃大模型及其原生 AI 终端,可实现跨端互联。 未来,无论是与阶跃大模型对话,或在其原生 AI 终端上,都不用打开 App, 一句话就能向其自有智能体派活 ,再转交阿宝办妥,完成跨应用、多任务执行。 据官方透露,合…

AI 点评 · AI生态打通,跨应用多任务执行,开启无需打开App的智能生活新场景。

IT之家
行业信号
7/17 09:13
打造开放可信普惠生态:AI 智能体互信互联互操作全球合作倡议发布

IT之家 7 月 17 日消息,2026 世界人工智能大会暨人工智能全球治理高级别会议主论坛今日在上海举行。会上,中国网信办会同有关方面正式提出《智能体互信互联互操作全球合作倡议》。 该倡议旨在释放智能体赋能可持续发展的潜力,防止形成智能鸿沟,凝聚各方共识,与全球伙伴共同打造开放、可信、安全、普惠的智能体生态。 智能体作为人工智能时代最具变革性的技术形态之一…

AI 点评 · 聚焦智能体生态的全球治理,推动开放互信,防范技术鸿沟,意义深远。

IT之家
设计 / 产品
7/17 08:46
网易智企携全新升级的一站式AI应用服务亮相WAIC 2026

36氪获悉,7月17日,在2026世界人工智能大会(WAIC 2026)上,网易智企携全新升级的一站式企业AI应用服务亮相,集中展示AI Agent编排、AI Coding、AI客服、AI私域助理、AI智能数据与AI Agent安全等企业级AI能力,围绕安全治理、组织协作与业务增长三大场景,呈现企业级AI应用实践。

AI 点评 · 展示企业级AI全栈能力,聚焦安全治理与业务增长三大场景,为行业提供可落地的AI应用标杆。

36氪
行业信号
7/17 07:25
阶跃终端首款智能体手机 STEPX Neo 现场实拍,Amoo 助手可与飞书等 App 深度联动

IT之家 7 月 17 日消息,2026 世界人工智能大会(WAIC 2026)今天在上海举办,IT之家第一时间来到阶跃展台,看到了阶跃终端首款智能体手机 STEPX Neo,下面为大家带来现场实拍: 从现场实拍可以看到,这台手机目前戴着橙黄色的保护壳, 运行智能体原生系统 Step AOS 。其桌面 UI 采用近年来较为流行的圆角矩形图标,部分第三方应用的…

AI 点评 · 智能体手机首次深度联动办公App,展示了AI原生系统的落地可能。

IT之家
设计 / 产品
7/17 03:26
Eval-core/evalcore

Snapshot testing for LLM apps and agents, built to run locally and block regressions in CI.

GitHub
设计 / 产品
7/17 03:26
Eval-core/evalcore

Snapshot testing for LLM apps and agents, built to run locally and block regressions in CI.

GitHub
设计 / 产品
7/16 21:21
kirodotdev/KiroCrew

A persistent workspace for development work that self-improves and continues beyond one session.

GitHub
模型 / Agent
7/16 19:29
Introducing Grok on Amazon Bedrock

This post covers what makes Grok 4.3 a great fit for agentic and enterprise workloads, how you access it through Amazon Bedrock, and how to use the capabilities most teams reach fo…

AI 点评 · Grok 4.3登陆AWS,为智能代理和企业场景带来新选择,看点在于云上部署的便捷性与性能优势。

AWS ML
论文 / 方法
7/16 17:54
Beyond Success Rate: Cost-Aware Evaluation of Offensive and Defensive Security Agents

Security-agent evaluations commonly measure peak offensive capability under generous inference budgets, emphasizing vulnerability discovery, exploit development, penetration testing, and CTF completio…

AI 点评 · 成本感知评估揭示安全AI实用门槛,突破传统成功率局限,更贴近真实攻防场景。

arXiv
论文 / 方法
7/16 17:45
AutoSynthesis: An agentic system for automated meta-analysis

Evidence synthesis is crucial for turning primary research into reliable knowledge for science, medicine, education, and policy. Yet, quantitative evidence synthesis remains largely manual and difficu…

AI 点评 · AI自动化元分析系统,极大提升科研效率,降低人工成本,推动循证决策发展。

arXiv
行业信号
7/16 15:38
Yes, you can now order DoorDash from the command line

DoorDash is opening a limited beta of dd-cli, a command-line tool that lets developers and AI agents search stores, build carts, and place orders from the terminal, marking another…

AI 点评 · 命令行点外卖,AI代理可直接下单,开启餐饮服务新交互方式。

TechCrunch
设计 / 产品
7/16 07:57
MemTensor/memmy-agent

🍙 A personal AI agent & local memory hub for all AI agents, gives every AI one shared, fully controlled memory and persistent context — all AI remember the sa…

GitHub
设计 / 产品
7/16 07:48
RoyZhao1991/LingShu

Apache-2.0, model-agnostic macOS agent: bring your own model to deliver verified code, docs, slides, and computer actions.

GitHub
Skill / 资源
7/16 00:33
Mermaid to Unicode box art (grok-mermaid)

Tool: Mermaid to Unicode box art (grok-mermaid) While exploring the codebase for the newly open-sourced Grok CLI coding agent I came across xai-grok-markdown/src/mermaid.rs , a "se…

AI 点评 · 将Mermaid图表转为Unicode字符画,适合终端环境,实用且有趣。

Simon Willison
模型 / Agent一手源
7/16 00:00
How Cars24 scales conversations and builds faster with OpenAI

Cars24 uses OpenAI-powered voice and chat agents to handle 1M+ monthly conversation minutes, recover 12% of lost leads, and bring agentic workflows to teams across the company.

AI 点评 · 用AI每月处理百万分钟对话,挽回12%流失客户,汽车电商实现业务流程自动化升级。

OpenAI
行业信号
7/15 23:16
豆包AI手机今年将发布多款机型

字节跳动联合中兴努比亚打造的首款AI智能体手机(“豆包AI智能体手机”)今年将有多款机型发布,其中一款将于2026世界人工智能大会期间亮相,其整体备货约20万台,首批备货10万台以内,截至发稿中兴方面对此消息暂无回应。(界面)

AI 点评 · 字节联手品牌进军硬件,多款AI手机布局显示生态野心,值得关注市场反应。

36氪
设计 / 产品
7/15 22:56
OpenAI 首款联名硬件:Codex Micro 键盘登场,灵活操控 AI 智能体

IT之家 7 月 16 日消息,OpenAI 今天(7 月 16 日)携手 Work Louder,合作推出 kbd-1.0-codex-micro 键盘,售价为 230 美元(IT之家注:现汇率约合 1560 元人民币)。 IT之家翻译产品官方描述如下: kbd-1.0-codex-micro 键盘采用 Work Louder 设计理念,实现 AI 智能体…

AI 点评 · OpenAI跨界做硬件,AI专用键盘或将重塑人机交互方式。

IT之家
设计 / 产品
7/15 22:46
gakonst/nanocodex

Building blocks for frontier OpenAI agents in Rust. Nanocodex empowers you with Codex-level performance anywhere.

GitHub
模型 / Agent
7/15 22:40
上传用户代码事件后,马斯克宣布开源 Grok Build 编程 AI 智能体工具

IT之家 7 月 16 日消息,马斯克旗下 SpaceXAI 公司昨日(7 月 15 日)宣布开源 Grok Build, 并将源代码发布至 GitHub 平台。 在官方博文中,SpaceXAI 表示: 开源发布源代码,是构建强大、可靠框架的最直接方法。用户可以阅读源代码,了解其从上下文构建到工具调用分发的完整工作原理。 开源也让框架更容易探索和扩展:如果用…

AI 点评 · 开源代码降低使用门槛,推动编程AI智能体技术快速迭代,值得开发者关注。

IT之家
行业信号
7/15 19:41
Amid hardware legal battle, OpenAI releases a $230 keyboard for Codex

OpenAI, which is in the middle of a legal battle with Apple over hardware trade theft allegations, just released a light-up keyboard designed to be paired with its agentic coding a…

AI 点评 · 硬件纠纷未平却推高价键盘,OpenAI跨界硬件野心与Codex生态绑定值得关注。

TechCrunch
行业信号
7/15 15:02
努比亚全球首款 AI 智能体手机局部外观公布,或采用横向镜组设计方案

IT之家 7 月 15 日消息,努比亚手机官方今日公布了旗下全球首款 AI 智能体手机的局部外观,新机将在 WAIC 2026 正式亮相。 预热图显示,努比亚全球首款 AI 智能体手机提供了一款淡粉配色,后盖中央印有“nubia”的字样,手机底部是扬声器开孔、USB-C 接口和 SIM 卡槽。遗憾的是,手机上端的摄像头模组被遮挡了,无法看见具体样式。 不过,…

IT之家
设计 / 产品
7/15 13:38
superdesigndev/treg

OpenRouter for agent tools. Join community here: https://discord.gg/6mQYYfFMAn

GitHub
行业信号
7/15 11:37
氪星晚报|LG新能源将为谷歌规模最大的“光储一体”项目供应电池;元宝与京东AI Agent正式打通小程序生态;日本散户持有美元净空头飙至2.79万亿日元,创2008年以来历史之最

大公司: 美团、青桔、哈啰共享单车调价 近期,美团单车、滴滴青桔、哈啰单车相继在北京等多个城市上调计费规则,三大平台不约而同地采取了“提高起步定价、拉长基础骑行时长”的组合策略:起步价从此前的1.5元/30分钟左右,普遍调整为1.88元至1.99元/60分钟。这成为共享单车行业近年来较大范围的一次集体调价。(金融时报) 瓜子二手车线下直卖场首店今日正式开业…

36氪
行业信号
7/15 09:50
支付宝不想做AI时代的配角

作者 | 王晗玉 编辑 | 张帆 支付宝首页调出AI界面,对话框取代了密密麻麻的小程序;用户对着“阿宝”说一句“找附近的奶茶优惠券”,周边门店的活动自动匹配好,核销下单一步完成。 最近,支付宝完成了上线22年来最大一次改版。 本月初,AI版支付宝“阿宝”正式开启全量公测,几乎同一时间,微信支付“AI专属卡”也在智能体WorkBuddy中落地。 进入2026年…

36氪
设计 / 产品
7/15 08:18
CyberSunil/LLMVault

An intentionally vulnerable OWASP LLM Top 10 training platform for AI Security, Prompt Injection, RAG Security, Agent Security, and GenAI penetration testing.

GitHub
设计 / 产品
7/15 02:49
0xsline/OpenChatCut

Open-source, local-first conversational AI video editor with a professional multi-track timeline, Agent Skills, MCP integration, and Remotion rendering.

GitHub
行业信号
7/15 02:40
36氪首发 | 前非夕科技核心业务合伙人创业,做垂域工业智能体,获数千万元种子轮融资

作者 | 乔钰杰 编辑 | 袁斯来 硬氪获悉,上海追知工程科技有限公司(以下简称“追知工科”)近日完成数千万元种子轮融资,由L2F光源创业者基金、尚融资本、一村资本联合投资。本轮融资将主要用于核心产品研发、团队建设及市场拓展。 追知工科成立于2024年2月,是一家 聚焦垂域工业智能体 的科技企业,同时也是上海交通大学成果转化企业、上海人工智能研究院战略孵化企…

36氪
Skill / 资源
7/14 18:44
Multi-agent social intelligence with Strands Agents and Amazon Bedrock

This post shows how Thrad.ai deployed a multi-agent system with Strands Agents and Amazon Bedrock AgentCore that automates the pipeline from prospect discovery through personalized…

AI 点评 · 多智能体协同与云平台结合,实现从线索挖掘到个性化服务的全流程自动化。

AWS ML
行业信号
7/14 11:21
氪星晚报 |智谱:完成配售新H股募资约314亿港元;荣耀与阿里将开展AI智能体终端合作;小米机器人首次实现汽车工厂柔性工件的长时作业

大公司: 中国神华:预计上半年净利润同比增长6.9%-21.1% 36氪获悉,中国神华公告,预计2026年上半年归属于上市公司股东的净利润为263亿元至298亿元,同比增长6.9%-21.1%。业绩变动主要系煤化工业务量及自有铁路、港口、航运业务量增加,带动相关业务利润同比增长。 中国人寿:预计上半年净利润同比增长约215%-235% 36氪获悉,中国人寿公…

36氪
行业信号
7/14 10:48
估值110亿美元,这家超级独角兽要帮初创公司从Day0走向全球

凌晨三点,一家刚成立不久的AI创业公司,可能已经在同时服务旧金山的客户、采购首尔的技术服务,并与拉各斯的合作伙伴签下合同。这家公司甚至还没有招到第一名全职财务人员,业务却已经跨越多个市场、币种和监管辖区。 AI正在让这样的创业路径成为可能。过去需要市场、运营、客服等一整套全球化团队才能完成的工作,现在借助智能体就能承担相当一部分。新一代初创企业不必再按照“先…

36氪
设计 / 产品
7/14 08:26
baldaworks/callee

Markdown-defined provider-backed agents and deterministic workflows for ACP runtimes.

GitHub
Skill / 资源
7/13 17:34
Building an agentic AI solution at Bluesight with Amazon Bedrock

In this post, we describe how Bluesight used two AWS engagements and Amazon Bedrock AgentCore to evolve from a single-product AI prototype to Prism, a unified agentic AI solution s…

AI 点评 · Bluesight借助亚马逊Bedrock将AI原型升级为统一代理系统,展示了企业级AI落地的实战路

AWS ML
设计 / 产品
7/13 08:57
try1004/FlowLens-AgentOps

Observable diagnostics, failure attribution, metrics, and no-key replay for multi-agent runtimes.

GitHub
设计 / 产品
7/13 05:51
cowboycodr/sitegeist

A visual benchmark testing whether leading coding agents repeat the same design patterns across 100 neutral website briefs.

GitHub
设计 / 产品
7/13 05:51
cowboycodr/sitegeist

A visual benchmark testing whether leading coding agents repeat the same design patterns across 100 neutral website briefs.

GitHub
行业信号
7/13 03:58
《智能体个人信息保护自律公约》发布,百度、腾讯、阿里、火山引擎等 31 家企业首批签署

IT之家 7 月 13 日消息,由中国互联网协会网民权益和个人信息保护工作委员会主办的 2026(第二十五届)中国互联网大会网民权益和个人信息保护论坛于 7 月 9 日在京举办。 中国互联网协会移动互联网工作委员会和中国互联网协会网民权益和个人信息保护工作委员联合发布《智能体个人信息保护自律公约》,现场来自 百度、腾讯、阿里、火山引擎 等 31 家互联网企业…

AI 点评 · 科技巨头联合签署,为AI智能体划清数据安全红线,行业自律迈出关键一步。

IT之家
设计 / 产品
7/13 01:02
iishyfishyy/operator-oss

Run many Claude Code (or Codex) sessions in parallel — across every project — from one screen. Local-first, git-worktree-isolated tasks, no API key.

GitHub
设计 / 产品
7/12 15:12
Meta 发布多模态推理模型 Muse Spark 1.1,强化 AI 智能体任务能力

IT之家 7 月 12 日消息,Meta 于 7 月 9 日正式发布适用于 AI 智能体的多模态推理模型 Muse Spark 1.1 版本,重点提升了模型在智能体任务中的规划、协同与执行能力,并增强了工具调用、代码开发、应用操作能力。 Meta 表示,Muse Spark 1.1 强化了多智能体协作机制,由主智能体负责收集信息、制定计划,再将任务拆分并分配…

AI 点评 · 多智能体协作机制是AI落地的关键突破,Meta这次强化了任务拆解与分工能力。

IT之家
设计 / 产品
7/12 14:34
manor-os/manor-ai

Source-available, self-hosted AI agent workspace for small businesses and lean teams — with BYOK, workflows, tools, and human approvals.

GitHub
设计 / 产品
7/12 11:44
S40911120/recensa

Self-hosted web viewer for Claude Code session transcripts — read, search, replay, and audit every session you have ever run

GitHub
模型 / Agent
7/11 11:35
智谱CEO唐杰发内部信:“GLM 时刻”和万亿俱乐部之后,什么是更重要的事

36氪独家获悉,7月11日,智谱创始人唐杰,在智谱发布了主题为《巨浪已来》的内部信。其中提到,智谱将不追求短期的应用变现,而是直指AGI的下一个高地:长程任务能力、完全自治的智能体系统、自我进化、极致安全治理。过去半年来,智谱收获了创立以来的高光时刻:市值较半年前上市初期涨了10倍,并在2026年 6月,跻身“万亿港元俱乐部”。

AI 点评 · 聚焦AGI核心突破而非短期变现,揭示智谱从估值飙升到技术深水区的战略转型。

36氪
行业信号
7/11 08:13
腾讯洽购Manus?知情人士:腾讯仍将保留少数股东地位

7月11日,有消息称腾讯正在洽谈成为通用AI Agent公司Manus的最大股东,据该消息,由腾讯牵头的中方资本组团以约20亿美元估值从Meta手中回购Manus的全部股权。记者向腾讯方面求证,截至发稿腾讯方面暂无回应。另有知情人士向记者透露,此次交易后,腾讯仍将保持少数股东地位,但不会控股。(南方都市报)

AI 点评 · 腾讯罕见以少数股东身份入局AI Agent,或意在布局生态而非控制,战略意图值得玩味。

36氪
模型 / Agent
7/11 01:24
9点1氪丨“国产存储第一股”长鑫科技公布承销团阵容;SK海力士登陆美股,上市首日大涨近13%;OpenAI推出ChatGPT智能体

今日热点导览 “全球首款智能体手机”已备货8万至10万台?知情人士:假的 百亿私募数量达142家,再次刷新历史纪录 三星李在镕拟于7月底赴美会晤英伟达黄仁勋 德国大众拟大裁员,最高或裁减12万个岗位 OpenAI高管层再现变动,首席运营官因病离职 TOP3大新闻 长鑫科技,承销团阵容公布 长鑫科技IPO进入发行倒计时,这家“国产存储第一股”背后的承销团阵容也…

AI 点评 · 国产存储芯片龙头IPO加速,承销团阵容披露凸显市场关注热度。

36氪
论文 / 方法
7/10 17:52
VEXAIoT: Autonomous IoT Vulnerability EXploitation using AI Agents

Internet of Things (IoT) systems are inherently vulnerable due to constrained hardware, outdated firmware, and insecure default configurations, creating a need for scalable and adaptive security testi…

AI 点评 · 用AI代理自动挖掘物联网漏洞,突破传统测试瓶颈,提升安全检测效率。

arXiv
论文 / 方法
7/10 16:54
Agora: Enhancing LLM Agent Reasoning Via Auction-Based Task Allocation

Enhancing the reasoning capabilities of large language model (LLM) agents requires effective orchestration of diverse expert models and tools. However, existing frameworks typically call APIs based on…

AI 点评 · 用拍卖机制分配任务,提升大模型推理效率,为多智能体协作提供新思路。

arXiv
行业信号
7/10 15:46
SK 集团会长崔泰源:AI 驱动内存需求呈指数级增长,未来五年产能翻倍仍难满足

IT之家 7 月 10 日消息,SK 海力士今日在纳斯达克挂牌交易其美国存托凭证(ADR)。随后,SK 集团会长崔泰源接受了彭博社与 CNBC 的采访。 IT之家注意到,崔泰源表示,在人工智能时代,内存行业已进入结构性增长阶段。 过去,内存的需求主要取决于人口数量或是智能手机和个人电脑的销量。 然而,随着 AI 智能体、推理过程中产生的键值缓存(KV Cac…

AI 点评 · AI引爆内存需求结构性变革,产能翻倍仍供不应求,预示行业进入长期高景气。

IT之家
设计 / 产品
7/10 15:32
Alisa0808/vox-director

Turn one topic into a finished Vox-style paper-collage explainer/ad video — automated end to end on Atlas Cloud + ffmpeg. An agent skill.

GitHub
Skill / 资源
7/10 15:28
Scaling agentic workflows with native case management in Amazon Quick Automate

In this post, we show you how to combine case management with agentic automation capabilities in Quick Automate. We introduce case management and explore the lifecycle of cases in…

AI 点评 · 结合案例管理与智能自动化,提升企业复杂任务处理效率,降低人工干预成本。

AWS ML
Skill / 资源
7/10 15:23
How KTern.AI built agentic AI for SAP on Amazon Bedrock AgentCore

Evolving from a traditional software as a service (SaaS) platform into a next-generation agentic AI platform meant orchestrating multiple specialized agents across long-running ent…

AI 点评 · KTern.AI用亚马逊Bedrock实现SAP多智能体协作,展示企业级AI落地新路径。

AWS ML
设计 / 产品
7/10 12:20
ShenSeanChen/waku-agent

Waku Waku! Waku agent is your personal AI agent, on your own laptop, in code you can read in an afternoon — harness + loop + memory + eval

GitHub
设计 / 产品
7/10 10:14
nexu-io/codex-slides

🎨 Open-source AI slide studio inside Codex: image-native decks, every slide a full visual canvas. ⚡ 10+ high-quality slides in ~4–5 minutes — Fast mode renders…

GitHub
设计 / 产品
7/10 04:13
AlephAITech/WorkBuddyGuide

A practical, open-source guide to mastering WorkBuddy through real-world workflows.开源的 WorkBuddy 实战蓝皮书:教程、真实工作流、Skills、MCP、自动化与多智能体实践。

GitHub
模型 / Agent
7/10 00:56
派早报:蔚来 ES8 大五座版正式上市等

OpenAI 发布 GPT-5.6 系列模型等,SpaceXAI 发布编程智能体模型 Grok 4.5。 查看全文

AI 点评 · 大模型竞速加剧,头部企业密集发布新品,技术迭代与市场格局值得追踪。

少数派
模型 / Agent
7/9 23:27
IT早报 0710:OpenAI 发布 GPT-5.6 系列模型;支付宝客服就花呗服务问题致歉;小米澎程首款 SUV 内外饰首秀;华为李文广回应友商智驾兜底...

“IT早报”时间,大家好,现在是 2026 年 7 月 10 日星期五,今天的重要科技资讯有: 1、OpenAI 最强 AI 模型:GPT-5.6 系列正式上线,纳德拉称微软 Copilot 同步接入 OpenAI 公司 7 月 10 日发布公告,宣布在 ChatGPT(聊天机器人)、Codex(主打编程 AI Agent,目前朝通用 Agent 方向)以及…

AI 点评 · OpenAI模型重大升级,微软深度整合,预示AI竞争进入新阶段。

IT之家
行业信号
7/9 23:15
奥尔特曼:GPT-5.6 Sol 是 OpenAI 最好 AI 模型,同等 / 更优性能下 Token 效率提高 54%

IT之家 7 月 10 日消息,在接受 CNBC 采访时,OpenAI 首席执行官萨姆 · 奥尔特曼(Sam Altman)表示,在 AI 智能体编程任务中, GPT-5.6 Sol 模型表现比市场主流竞争模型“一样好,甚至更好”,但 Tokens 效率提高 54%。 IT之家注:原文中并未具体指名市场主流竞争模型,不过鉴于 GPT-5.6 Sol 模型的定…

AI 点评 · 性能飞跃与成本降低同步实现,这项突破将加速AI应用落地。

IT之家
模型 / Agent
7/9 22:57
OpenAI 推出 ChatGPT Work 智能体:GPT-5.6 支持,可驾驭长时间、多步骤任务

IT之家 7 月 10 日消息,OpenAI 今天(7 月 10 日)发布博文,在宣布推出 GPT-5.6 系列 AI 模型的同时,还推出全新的 ChatGPT Work 智能体, 由 GPT-5.6 提供支持,定位为可承担长时、多步骤任务的智能体。 在博文中,OpenAI 披露了 Codex 的现有使用规模:官方数据显示,Codex 每周用户数已超过 50…

AI 点评 · GPT-5.6加持的Work智能体,标志着AI从对话迈向长周期复杂任务自主执行。

IT之家
模型 / Agent
7/9 22:46
OpenAI 最强 AI 模型:GPT-5.6 系列正式上线,纳德拉称微软 Copilot 同步接入

IT之家 7 月 10 日消息,OpenAI 公司今天(7 月 10 日)发布公告, 宣布在 ChatGPT(聊天机器人)、Codex(主打编程 AI Agent,目前朝通用 Agent 方向)以及 API 中上线 GPT-5.6 系列模型。 在模型方面,IT之家援引博文介绍,OpenAI 本次共发布 3 档模型: 旗舰版 Sol(太阳):每 100 万 T…

AI 点评 · 微软同步接入,意味着AI竞争格局突变,企业级应用迎来新拐点。

IT之家
行业信号
7/9 22:08
An AI agent startup just let its agent run its $100M fundraise

Lyzr, a startup that builds AI agents for enterprises, used its own AI agent to raise a $100 million round — proof, evidently, that the product actually works.

AI 点评 · AI代理成功完成融资,证明企业级产品真实可用,开创行业先河。

TechCrunch
行业信号
7/9 19:40
Meta enters the crowded AI coding battle with Muse Spark 1.1

Meta's pitch to users is Spark's ability to handle large agentic workloads, fix bugs, and help with large code migrations — the kind of automation that enterprises are increasingly…

AI 点评 · Meta携Spark 1.1切入企业级AI编码自动化,专注大型代码迁移和修复,展现巨头竞争新方向。

TechCrunch
设计 / 产品
7/9 16:24
Introducing Muse Spark 1.1

Introducing Muse Spark 1.1 Following Muse Spark in April , here's Muse Spark 1.1 - the first Spark model to offer an API. Meta claim significant improvements in agentic tool callin…

Simon Willison
模型 / Agent
7/9 10:00
ChatGPT is now a partner for your most ambitious work

ChatGPT Work is an agent that can take action across your apps and files, stay with a project for hours if needed, and turn a goal into finished work.

AI 点评 · ChatGPT从对话助手跃升为跨应用自主执行任务的智能代理,标志AI进入主动工作阶段。

OpenAI
设计 / 产品
7/9 06:46
UiPath/coder_eval

Evaluate & benchmark AI coding agents and Claude Code skills — sandboxed, reproducible YAML eval suites for Claude Code, Codex & Gemini, with A/B experiments an…

GitHub
模型 / Agent
7/8 22:51
马斯克 SpaceXAI 首个编程智能体模型 Grok 4.5 发布:与 Cursor 联合训练,效率翻倍价格减半

IT之家 7 月 9 日消息,SpaceXAI 今日正式发布了其 Grok 4.5 模型,这是该公司首个专门针对编程和智能体任务训练的模型。 据介绍,该模型由 SpaceXAI 与 Cursor 联合完成训练,在提供前沿智能水平的同时,兼具领先的速度与成本效率。马斯克将其称为“Opus 级模型”。 Grok 4.5 面向真实工程场景设计,擅长处理大型代码库以…

AI 点评 · 编程智能体成本砍半效率翻倍,马斯克联手Cursor的定价策略才真值得行业关注。

IT之家
设计 / 产品
7/8 19:28
xinhuangcs/agentmaker

A general-purpose Python framework for building LLM agents and multi-agent systems. "Four lines of code, an agent with memory."

GitHub
Skill / 资源
7/8 17:16
Data for Agents

AI 点评 · 数据驱动智能体的新范式,预示AI自主决策能力将迎来关键突破。

HuggingFace Blog
论文 / 方法一手源
7/8 16:00
Flint: A visualization language for the AI era

Short chart specifications are easy to write, but often produce uninspiring results. Flint is an open-source visualization language that offers a middle path, letting AI agents cre…

可信度 88交叉信源 1
Microsoft Research
设计 / 产品
7/8 13:02
BitMiracle-AI/Dormice

The SQLite of agent sandboxes — self-hosted, E2B-compatible. One machine, sandboxes that live forever, idle costs nothing.

GitHub
设计 / 产品
7/8 12:54
TencentCloud/Octop

A smarter, self-hosted AI assistant — multi-user, multi-agent.

GitHub
设计 / 产品
7/8 08:05
地产AI进入落地战,深度智联又甩出一张牌

7月7日,易居(中国)控股有限公司董事局主席、总裁周忻再一次来到台前,给公司的AI产品站台,推出其核心战略产品“地产模数通——企业专属大模型一体机”。同时,克而瑞地产AI分析师“小瑞”正式上岗。 不到两个月前,易居旗下的深度智联刚刚发布了全球首个房地产经纪人智能体“易居·小新”,用中立无佣模式替代传统中介。迹象显示,深度智联在加速AI产品落地。 周忻说,地产…

36氪
行业信号
7/8 00:00
「德睿智药」获5200万美元B轮融资,AI设计的减肥药已进入3期临床|36氪首发

文|胡香赟 编辑|海若镜 36 氪获悉,德睿智药近期已完成 5200 万美元B轮融资,投资方包括头部人民币和美元基金,凯乘资本为独家财务顾问。募集资金将用于AI制药引擎Molecule Arts Platform(MAP)升级迭代,完善其多智能体(Multi-Agent)协同体系与临床数据闭环(Clinical Data-in-the-Loop),以及推进自…

36氪
Skill / 资源
7/7 16:51
Build a serverless image editing agent with Amazon Bedrock AgentCore harness

This post walks through building a serverless image editor where users upload a photo, describe an edit in plain English, and receive the result in seconds. The agent runs on Agent…

AI 点评 · 无服务器图像编辑结合自然语言指令,大幅降低AI应用门槛,展示Bedrock AgentCore的实用

AWS ML
设计 / 产品
7/7 11:10
EXXETA/exxperts

Local-first AI agents with governed, approval-gated memory. Any model provider; MCP tools and web search built in. Nothing remembered without your say-so, nothi…

GitHub
设计 / 产品
7/7 08:38
caseclose/cma-harness

Cognitive-structured Multimodal Agent (CMA-Harness): a memory-centric agent for long-horizon multimodal understanding, generation, and editing — externalizing v…

GitHub
设计 / 产品
7/7 06:23
"龙虾"为什么这么火?OpenClaw登顶GitHub后,AI Agent时代真的来了?

GitHubStar-history 最近Openclaw以25.2万星标,超越Meta的React登顶GitHub开源项目历史第一! 要知道React是Facebook(现改名Meta)打造的经典前端框架,过去十余年间,互联网上绝大多数我们熟知的网站与App,底层技术架构皆由它构筑。Openclaw官方更是高调发文嘲讽Meta“我们在迭代创新,而你只在办会…

36氪
设计 / 产品
7/7 02:44
frenzymath/Danus

Orchestrating Mathematical Reasoning Agents with Fact-Graph Memory

GitHub
设计 / 产品
7/6 22:35
Sahir619/fable-method

The Fable Workflow: how Claude Fable 5 worked, distilled into skills any model can run, with the eval that keeps it honest. Think / act / prove.

GitHub
设计 / 产品
7/6 16:18
ronak-create/FableCut

Zero-dependency browser video editor that AI agents can drive — JSON timeline, MCP + REST, live-reloading UI

GitHub
设计 / 产品
7/6 16:18
ronak-create/FableCut

Zero-dependency browser video editor that AI agents can drive — JSON timeline, MCP + REST, live-reloading UI

GitHub
设计 / 产品
7/6 14:56
nexu-io/motion-anything

✨ The agentic motion layer — an open-source, chat-native motion engine. Describe the feeling; your AI ships the animation.

GitHub
设计 / 产品
7/6 14:56
nexu-io/motion-anything

✨ The agentic motion layer — an open-source, chat-native motion engine. Describe the feeling; your AI ships the animation.

GitHub
行业信号
7/6 07:53
获DCM Ventures投资数百万美元,APTSell希望成为AI版的首席销售官|涌现新项目

文|吴思瑾 编辑|邓咏仪 01 一句话介绍 北京治真治合科技有限公司成立于2024年,旗下产品「APTSell」(AI Power To Sales��希望成为AI版的CSO (Chief Sales Officer,首席销售官)。 简单来说,APTSell是一个组合式Agent,通过整合与可视化销售全流程数据,生成管理决策和执行建议,以期正向促进销售效率和…

36氪
设计 / 产品
7/5 23:40
派早报:阿里禁用 Claude 模型

阿里禁用 Claude 模型 索尼调整计划,2028 年前发售游戏可继续生产光盘 千问、豆包将下线智能体功能 Android 反垄断案欧洲终审败诉 混动车、商用纯电车将不再免征车船税 电商法修正案征求意见 看看就行的小道消息 少数派的近期动态 你可能错过的好文章 查看全文

少数派
设计 / 产品
7/5 23:25
oxbshw/watch-skill

Video understanding and self-verification for AI agents. Turn videos, streams, and agent screen recordings into searchable, timestamped evidence—then use THE LO…

GitHub
设计 / 产品
7/5 15:48
Bike4Mind/bike4mind

The open-core AI workbench — notebooks, agents, RAG, voice, and images across any model: OpenAI, Anthropic, Google, xAI, or local via Ollama/vLLM. BSL 1.1, aut…

GitHub
设计 / 产品
7/5 09:23
ContextJet-ai/awesome-llm-observability

50+ curated LLM observability tools PLUS 26 Agent Skills (several with runnable, unit-tested scripts) to build, evaluate, debug, secure & monitor reliable LLM a…

GitHub
设计 / 产品
7/5 06:30
simonlin1212/Vibe-Research

Vibe-Research: Your Personal Trading Research Agent · A股/美股/港股 的个人投研 Agent:每日复盘、资讯雷达、个股数据、板块中心、我的持仓、研究记录。Vibe-Research 把数据和功能配齐,由你自己的 AI 驱动投资研究。

GitHub
设计 / 产品
7/5 05:38
zhiweio/EagleRAG

Search knowledge by what documents mean and how they look — not one or the other.

GitHub
设计 / 产品
7/4 15:28
SmileLikeYe/agent-chief

Attention is your scarcest resource. Chief is the local-first layer that guards it — turning every agent, alert, and feed into one honest call: interrupt, or no…

GitHub
设计 / 产品
7/3 21:09
voly-codes/voly

Control plane for AI coding agents: route tasks, reduce token spend, run multi-agent workflows, fallback executors, and track cost per task.

GitHub
设计 / 产品
7/3 14:16
豆包智能体功能将于 7 月 15 日下线,官方建议提前完成备份

IT之家 7 月 3 日消息,豆包今晚发布《豆包智能体功能下线通知》,称由于产品功能调整, 智能体功能将于 2026 年 7 月 15 日下线 。 《通知》显示,该功能下线后,用户仍可在一段时间内查看并自行保存智能体信息及历史对话数据。2026 年 10 月 15 日后,豆包将根据《隐私政策》对智能体相关数据进行处理, 后续将无法在豆包内查看或恢复 。如有重…

AI 点评 · 产品功能调整背后,需关注用户数据迁移与隐私政策变化对AI服务稳定性的影响。

IT之家
设计 / 产品
7/3 09:22
ai4s-research/open-science

Open Science Desktop — local-first, model-agnostic AI research workbench for macOS, Windows & Linux. Open-source Claude Science desktop alternative built on Tau…

GitHub
设计 / 产品
7/3 02:50
aipoch/open-science

Open Science is an open-source, local-first, model-agnostic AI research workbench for scientific discovery.

GitHub
行业信号
7/2 23:22
Meta CEO 马克 · 扎克伯格:AI 智能体技术发展得比我想象要慢

IT之家 7 月 3 日消息,据《商业内幕》今天报道,Meta 首席执行官马克 · 扎克伯格在上周四的一场内部全员会中表示,公司仍在努力实现“超级智能”(Superintelligence),但目前还需要投入更多时间和精力。 据两位参会人士透露,扎克伯格表示,Meta 正在向人工智能领域投入大量资源, 但 AI Agent(IT之家注:AI 智能体)技术的发…

AI 点评 · 行业领袖坦言进度不及预期,揭示AI智能体落地瓶颈,值得关注其实际挑战与未来方向。

IT之家
Skill / 资源
7/2 19:33
llm-coding-agent 0.1a0

Release: llm-coding-agent 0.1a0 Another Fable 5 experiment. Now that my LLM library has evolved into more of an agent framework it's time to see what a simple coding agent would lo…

AI 点评 · 首个开源LLM编程代理框架,简化代码生成与迭代流程,开发者可快速上手实验。

Simon Willison
论文 / 方法
7/2 17:59
Distributed Attacks in Persistent-State AI Control

As AI coding agents become more autonomous, they increasingly ship code iteratively, with the codebase persisting across sessions. This persistence creates a new attack surface: a misaligned or prompt…

arXiv
论文 / 方法
7/2 17:55
Controllable Sim Agents with Behavior Latents

Realistic traffic simulation requires agents that imitate logged behavior and can also be steered along interpretable axes. Such controllability enables engineers to isolate variables, reproduce speci…

arXiv
设计 / 产品
7/2 17:53
elder-plinius/T3MP3ST

autonomous red teaming platform; multi-agent offensive-security meta-harness

GitHub
设计 / 产品
7/2 15:55
NimaChu/agent-wiki

Zero-cost, beginner-friendly local Markdown knowledge base for AI agents. Capture sources, preserve evidence and images, synthesize wiki pages, search, lint, an…

GitHub
设计 / 产品
7/2 15:55
NimaChu/my-wiki

Local-first AI knowledge app and Agent Skill with evidence-backed Wiki, an interactive knowledge universe, Viki Q&A, and shareable knowledge galaxies.

GitHub
设计 / 产品
7/2 15:55
NimaChu/my-wiki-skill

Agent Skill for building evidence-backed Markdown knowledge bases with zero-cost setup, image-aware capture, automatic wiki maintenance, and an interactive know…

GitHub
行业信号
7/2 13:22
谁能想到,系统流「爽文」最先被AI Agent实现了

撰文|深海 网文里的“系统流”,被拍成了职场短剧 千禧年初的网文圈,有三大经典题材在爽文届立于不败之地:无限流、快穿流、系统流。 这三大爽文战神体横空出世时,对IP界几乎是降维打击。当传统小说还在费劲搭世界观、铺人物成长弧光时,系统流已经绕过漫长的发育过程,直接把爽感推到最大。系统,这个堪称bug的存在,无论主角进入什么样的世界副本,面对不同的任务、危机和奖…

36氪
Skill / 资源
7/1 18:03
Structured memory filtering with metadata in AgentCore Memory

In this post, you will learn how metadata works across configuration, ingestion, and retrieval, explore enterprise use cases including multi-agent and multi-tenant architectures, a…

AI 点评 · 元数据过滤让AI记忆更精准,支撑多智能体架构,推动企业级应用。

AWS ML
Skill / 资源
7/1 17:53
How Inscribe uses Amazon Bedrock to stop document fraud in seconds

In this post, you will learn how Inscribe developed an agentic AI system using Amazon Bedrock that reasons across documents the way an expert fraud analyst would. With this new age…

AI 点评 · 利用亚马逊Bedrock的智能体AI,数秒内模拟专家分析文档,革新防伪效率。

AWS ML
设计 / 产品
7/1 07:57
xuzhougeng/wisp-science

Open-source, local-first desktop AI research workbench for scientific computing with Python/R, MCP bioinformatics tools, SSH/WSL/GPU runtimes, and OpenAI/Anthro…

GitHub
设计 / 产品
7/1 04:28
isjiamu/gzh-design-skill

把 Markdown 一键排成可直接粘进公众号编辑器的精致 HTML —— 6 套精选主题 + 主题生成器 + 双关卡校验。An AI-agent skill that turns Markdown into paste-ready WeChat article HTML.

GitHub
模型 / Agent
6/30 18:00
Anthropic launches Claude Sonnet 5 as a cheaper way to run agents

Anthropic’s Claude Sonnet 5 brings stronger agentic capabilities, lower pricing, and improved safety, positioning the model as a cheaper alternative to Opus, GPT-5.5, and Gemini Pr…

AI 点评 · 低价推出强智能体能力,性价比对标顶级模型,AI行业竞争白热化。

TechCrunch
论文 / 方法
6/30 17:53
Generative Skill Composition for LLM Agents

Recent LLM agents benefit from skills for solving complex tasks. Skills encapsulate modular packages of procedural knowledge and instructions for performing specialized tasks, such as setting up a san…

arXiv
行业信号
6/30 17:52
Acti puts AI agents directly into your smartphone keyboard

Acti is betting the smartphone keyboard is the next home for AI assistants. The startup's new keyboard for iOS and Android works across apps and lets users create custom AI-powered…

AI 点评 · 用键盘直接调用AI智能体,打破应用壁垒,或成手机助手新入口。

TechCrunch
论文 / 方法
6/30 17:39
AxDafny: Agentic Verified Code Generation in Dafny

We study agentic code generation in Dafny, where a model must generate both executable code and the proof artifacts for verification. We present AxDafny, a verifier-guided repair framework that iterat…

arXiv
论文 / 方法一手源
6/30 16:50
SkillOpt: Agent skills as trainable parameters

AI agents often fail because their instructions, or skills, are manually modified with no guarantee of improvement. Learn how SkillOpt turns skill editing into a training process,…

可信度 88交叉信源 1
Microsoft Research
行业信号
6/30 15:27
华为官宣全球首个商用多模态文旅大模型规模化应用

IT之家 6 月 30 日消息,华为中国宣布,2026 年 6 月 29 日,全球首个商用多模态文旅大模型 ——“博观文旅大模型”在西安规模应用。截至今年 3 月, “博观”支撑开发的 AI 伴游智能体已覆盖超 400 万用户 。其打造的非遗数字 IP,衍生产品销售超 200 万。 IT之家查询获悉,陕文投与华为等于 2025 年 9 月联合开发的“博观文旅…

IT之家
Skill / 资源
6/30 15:10
shot-scraper 1.10

Release: shot-scraper 1.10 The big new feature is shot-scraper video storyboard.yml , described in detail in Have your agent record video demos of its work with shot-scraper video…

Simon Willison
设计 / 产品
6/30 14:34
Orkas-AI/Orkas-VideoStudio

Turn your coding agent into a video studio: describe a video in plain language, and your agent writes the timeline and produces the file.

GitHub
设计 / 产品
6/30 06:42
runvendo/vendo

Embedded agents your customers use to automate work, build views, and connect their tools.

GitHub
设计 / 产品
6/29 23:13
redevops-io/context-runtime

Context Runtime — a database query planner for LLM context. Decides what a model sees before it answers; plans it, runs it through reused substrate, and learns…

GitHub
论文 / 方法
6/29 20:00
Xiaomi-GUI-0 Technical Report

Graphical user interface (GUI) agents build on vision-language models to complete user tasks end-to-end in real applications through interface actions such as tapping, swiping, text entry, and navigat…

HuggingFace Papers
行业信号
6/29 18:00
AI agents are not your “coworkers”

This story originally appeared in The Algorithm, our weekly newsletter on AI. To get stories like this in your inbox first, sign up here. Imagine coming in to work to learn that a…

AI 点评 · 点明AI助手定位偏差,揭示行业对“AI同事”概念的过度浪漫化。

MIT Tech Review
论文 / 方法
6/29 17:58
Self-Evolving World Models for LLM Agent Planning

World models offer a principled way to equip long-horizon LLM agents with foresight: predictions of action consequences before execution. However, unreliable foresight can be ignored, misused, or even…

arXiv
Skill / 资源
6/29 17:25
Debugging production agents with Amazon Bedrock AgentCore Observability

In this post, you learn how to debug production agent failures using built-in observability capabilities. We walk through common failure patterns, show how to analyze agent behavio…

AI 点评 · 用内置可观测性调试生产级AI代理,解决落地部署中的常见故障分析难题。

AWS ML
论文 / 方法
6/29 17:14
Attractor States Emerge in Multi-Turn LLM Conversations

Large language models (LLMs) are increasingly used in open-ended multi-agent settings, but the long-run dynamics of model--model interaction remain poorly understood. We study whether open-ended LLM d…

arXiv
Skill / 资源
6/29 16:17
Ornith-1.0: Self-Scaffolding LLMs for Agentic Coding

Ornith-1.0: Self-Scaffolding LLMs for Agentic Coding This is an interesting new open weights (MIT licensed) model, the first model release from DeepReinforce. [...] with variants i…

AI 点评 · 首个自搭建代码智能体模型开源,MIT许可降低门槛,或重塑AI编程工具生态。

Simon Willison
设计 / 产品
6/29 14:05
MHW888888/aegisloop

Open-source local policy, recovery, and audit layer for explicit Codex execution.

GitHub
行业信号
6/29 03:43
36氪首发 | URTOPIA联创做了款智能指环,众筹已破千万元

作者 | 张子怡 编辑 | 袁斯来 近日,AI可穿戴品牌AIVELA宣布完成数百万美元首轮融资。本轮融资由线性资本领投,锋领资本跟投,智能电助力自行车品牌URTOPIA等产业方共同加注。 本轮融资将主要用于下一代AI可穿戴产品研发、健康数据与AI Agent能力建设、全球市场拓展以及核心团队扩张。AIVELA将以智能指环、智能手链等贴身可穿戴产品为起点,面向…

36氪
Skill / 资源
6/28 21:57
Quoting Jon Udell

Human Agent in the loop I dislike the phrase “human in the loop” because it cedes authority to the machines. Let’s flip the narrative. It’s our loop, we work the same way we always…

AI 点评 · 观点颠覆:主张人类掌控AI工具,而非被机器主导,重新定义人机协作的主动权。

Simon Willison
设计 / 产品
6/28 20:01
kaderkck/hewn-forge

HEWN 2.0 2026: AI Output Router for Precision Summaries & Polished Code

GitHub
设计 / 产品
6/28 20:01
kaderkck/hewn-forge

HEWN 2.0 2026: AI Output Router for Precision Summaries & Polished Code

GitHub
设计 / 产品
6/28 15:20
mingchen666/Reviva

Local-first AI learning workspace — ask, note, review and create around your own materials. Wiki KB, Agents, Skills, creation tools.AI 学习工作台,围绕你的资料完成问答、笔记、复习和…

GitHub
设计 / 产品
6/28 03:39
deer-flow/llm-space

A desktop app to prototype agent ideas, inspect every harness step, replay failures, and evaluate performance, all in one place. Local-first, cloud-ready for ma…

GitHub
行业信号
6/27 20:34
Show HN: Adrafinil – keep a lid-closed Mac awake only while agents work

A month ago there was a wave of posts and tweets about engineers walking around cafes and parks with their MacBooks propped half-open, as fully closing the lid forces sleep that st…

AI 点评 · 针对macOS使用痛点,用AI智能管理开盖唤醒,提升开发者外接设备时的工作效率。

Hacker News
设计 / 产品
6/27 20:11
owainlewis/push

A secure, stable, and lightweight alternative to OpenClaw and Hermes.

GitHub
论文 / 方法
6/27 20:00
Hierarchical Experimentalist Agents

Large language models (LLMs) are increasingly used to take actions in the real world and support human decision-making, yet most agents rely on parametric knowledge, fixed post-training data, retrieva…

HuggingFace Papers
设计 / 产品
6/27 13:01
simonlin1212/investment-news

为 A股投资者打造的全球产业链资讯看板 · 12 大赛道一一对应 A股板块(半导体/AI/机器人/新能源车…),覆盖 100+ 权威源,用你自己的大模型每日提炼中文「今日要点」+ 翻译 · 全程本地、零 API key · Local AI news dashboard tracking the global indu…

GitHub
Skill / 资源
6/27 11:21
Using Local Coding Agents

Using Open-Weight Models in Local Coding Harnesses as an Alternative to Claude Code and Codex Subscriptions

AI 点评 · 本地化开源模型替代付费服务,降低AI编码成本与依赖。

可信度 74交叉信源 1
Sebastian Raschka
设计 / 产品
6/26 23:10
市场监管总局:加快智能体、具身智能等前沿技术领域标准制定速度

IT之家 6 月 27 日消息,据央视新闻 6 月 25 日报道,市场监管总局正会同相关部门,加快智能体等前沿技术领域标准制定速度,动态完善适配产业发展的人工智能国家标准矩阵。 报道称,目前正在抓紧制定的国家标准,除智能体外,还有 具身智能、世界模型、本体模型 等前沿技术标准,算力基础设施、高质量数据集、仿真测试平台、深度学习编译器、开源模型平台等底座类标准…

AI 点评 · 政策加速标准制定,将推动智能体与具身智能产业规范化发展,抢占技术制高点。

IT之家
Skill / 资源
6/26 17:58
Incident Report: CVE-2026-LGTM

Incident Report: CVE-2026-LGTM Spectacular hypothetical incident report by Andrew Nesbitt. Day 2, 16:00 UTC --- Two AI review agents from competing vendors, both attached to a down…

AI 点评 · 虚构安全事件揭示AI代理间冲突风险,警示未来协作需防漏洞。

Simon Willison
行业信号
6/26 16:40
Show HN: Smart model routing directly in Claude, Codex and Cursor

We built a model router that plugs into coding agents (e.g. Claude Code, Codex, Cursor, etc.) and intelligently sends requests to the best model to serve them. Here's a quick demo…

AI 点评 · 让编程助手自动匹配最优模型,大幅提升代码生成效率与成本控制。

Hacker News
Skill / 资源
6/25 22:28
AI and Liability

AI and Liability Bruce Schneier and Nathan Sanders on the recent German ruling that Google be held liable for errors introduced in their AI overviews: AI agents are agents of the p…

AI 点评 · AI责任判定首案,德国法院裁定谷歌为AI错误担责,确立行业先例。

Simon Willison
论文 / 方法
6/25 02:00
How agents are transforming work

A new OpenAI research paper shows how AI agents are transforming work, enabling longer, more complex tasks and expanding productivity across roles.

OpenAI
Skill / 资源
6/24 18:20
Build a healthcare appointment agent with Amazon Nova 2 Sonic

In this post, you will learn how to build a voice agent that handles appointment reminder conversations using Amazon Nova 2 Sonic and Amazon Bedrock AgentCore. The agent authentica…

AI 点评 · 亚马逊推出低成本语音预约系统,展示AI在医疗场景的落地新范式。

AWS ML
Skill / 资源
6/24 16:56
How Loka Built a Natural, Low-Latency Voice Agent with Amazon Nova 2 Sonic

In this post, we demonstrate the architecture and approach Loka used to solve a common frustration: robotic, slow voice assistants that cause customers to hang up, damaging brand r…

AI 点评 · Loka用亚马逊新模型打造自然流畅语音代理,解决机器人语音痛点,技术方案值得借鉴。

AWS ML
设计 / 产品
6/24 15:25
ZeKaiNie/universal-examprep-skill

Last-night exam-cram coach as a Claude Agent Skill: turns your slides, notes and past papers into a chaptered knowledge base + quiz bank, teaches only what's in…

GitHub
行业信号
6/24 10:06
完成数亿元新融资,影眸科技 Hyper3D 让 3D 生成进入“思考时代”丨36氪首发

文|王欣逸 编辑|张雨忻 2026 年开年来,3D 生成模型赛道相当热闹。 今年第一季度,影眸科技发布首个 3D 编辑模型 Rodin Gen-2 Edit,让 AI 3D 模型第一次可编辑;今年 6 月,VAST 官宣了新一轮融资,Meshy 也紧随其后,宣称自己发布了全球首款 3D AI Agent。 近日,影眸科技——这支扎根学术圈、创业早、年轻的 3…

36氪
设计 / 产品
6/24 08:10
benchflow-ai/awesome-evals

A curated, non-BS library of the best resources for building and evaluating AI agents — papers, blogs, talks, tools, benchmarks. Maintained by BenchFlow.

GitHub
设计 / 产品
6/24 07:12
raiyanyahya/llmaker

Selfhost modern LLM stacks. Run the whole fleet from your terminal

GitHub
设计 / 产品
6/24 03:23
fancyboi999/open-tag

Open-source, self-hostable alternative to Claude Tag — a Slack-style workspace where your team and its AI agents (Claude Code, Codex, GitHub Copilot, and more)…

GitHub
设计 / 产品
6/23 23:23
英伟达发布BioNeMo Agent工具包

当地时间6月23日,英伟达宣布推出NVIDIA BioNeMo Agent Toolkit,该工具包包含英伟达超过十年的生命科学库、工具和开放模型,使AI智能体、科学家和实验室能够通过收集证据、跨研究结果进行推理、运行计算实验以及推荐下一步最佳行动来协同工作,从而加速科学发现。(界面)

AI 点评 · 英伟达BioNeMo将十年生命科学积累与AI智能体结合,有望大幅加速药物研发和科学实验。

36氪
设计 / 产品
6/23 17:19
usedotai/dot-loom

Provider-pluggable orchestration runtime for multi-model AI inference. ( Sakana Fugu style )

GitHub
设计 / 产品
6/23 17:09
Reyzowter/Hello-Agents

🤖 Building AI Agent Systems from Scratch — A comprehensive, practical tutorial from fundamentals to production-grade multi-agent applications

GitHub
设计 / 产品
6/23 17:09
Reyzowter/Hello-Agents

🤖 Building AI Agent Systems from Scratch — A comprehensive, practical tutorial from fundamentals to production-grade multi-agent applications

GitHub
行业信号
6/23 08:31
Show HN: Neural Particle Automata

Neural CAs model self-organizing pattern formation on grids. Now the grid is gone. Each cell is an agentic particle that can move freely in space and change its state. While each p…

Hacker News
设计 / 产品
6/23 08:16
有智青年挑战赛暨全国AI+场景应用大赛决赛收官!在WAVES 2026的舞台上,挖掘中国下一代AI独角

有智青年挑战赛暨全国AI+场景应用大赛决赛在WAVES新浪潮大会期间举行,汇聚多支青年团队围绕AI与数字经济前沿场景展开角逐,展现青年创业者的技术探索与落地能力。 2026年,AI全面进入“行动者”时代。当大模型、智能体、具身智能从实验室的技术概念,全面走进千行百业的产业落地场景,AI工具的平民化与普及化,依托开源生态、轻量化开发工具与普惠算力,不断压低创新…

36氪
设计 / 产品
6/23 07:11
vancyland/DataClaw0

DataClaw: Agentic Tailoring Multimodal Data from Raw Streams — coming soon (code, weights, dataset & DataClaw-val upon acceptance).

GitHub
设计 / 产品
6/23 07:11
vancyland/DataClaw0

DataClaw: Agentic Tailoring Multimodal Data from Raw Streams — coming soon (code, weights, dataset & DataClaw-val upon acceptance).

GitHub
行业信号
6/22 20:53
The AI world is getting ‘loopy’

The loop takes agentic AI a step further by authorizing a swarm of agents to work continuously in the background, endlessly.

AI 点评 · 自主代理集群持续后台运行,突破传统AI单次交互模式,预示自动化新纪元。

TechCrunch
设计 / 产品
6/22 13:43
Johell1NS/browser-search

A skill for AI agents: search the web with SearXNG, browse with Camofox, bypass protections with CloakBrowser. Anti-hallucination by design. Self-hosted, free,…

GitHub
设计 / 产品
6/22 05:55
NotASithLord/peerd

The first AI agent harness native to the browser. A browser extension that runs a full agent loop where you already work: it drives your tabs, spins up sandboxe…

GitHub
行业信号
6/21 23:34
中信建投:国产模型加速迭代,算力景气度持续

36氪获悉,中信建投研报称,国内模型持续迭代,GLM-5.2、Kimi K2.7 Code强化1M上下文、长程Agent、Agentic Coding和真实工程交付能力,推动国产模型从通用问答转向开发者工具和企业级工作流。Kimi补强国际化运营能力,DeepSeek融资强化头部模型产业化预期,微信AI灰度测试则显示AI入口正从独立App走向超级应用生态,有望…

AI 点评 · 国产模型转向企业级应用,算力需求确定性增强,产业链景气度有望持续。

36氪
Skill / 资源
6/21 22:01
Temporary Cloudflare Accounts for AI agents

Temporary Cloudflare Accounts for AI agents The announcement says this is "for AI agents" but (as is pretty common these days) the AI hook isn't really necessary, this is an intere…

AI 点评 · 为AI代理提供临时账户,简化安全访问,降低管理成本,是云服务与自动化结合的新探索。

Simon Willison
设计 / 产品
6/21 20:44
redevops-io/redevops-rag

Hybrid RAG (DuckDB vector + BM25 + RRF + recency/keyword priors + optional cross-encoder rerank) as an installable library + CLI.

GitHub
论文 / 方法
6/21 20:00
Tmax: A simple recipe for terminal agents

Terminal-using agents have quickly become the most popular downstream application of language models (LMs). Despite their prevalence, relatively little academic work has examined RL-based training of…

HuggingFace Papers
论文 / 方法
6/21 20:00
Training Open Models for Agentic Phone Use

Phones are becoming an important execution surface for general-purpose agents, but training open models for reliable phone use remains difficult because the environment that matters at deployment, rea…

HuggingFace Papers
论文 / 方法
6/21 20:00
Self-Compacting Language Model Agents

Long agent traces composed of chains of thought and tool calls accumulate stale content that anchor subsequent generations, and eventually outgrow the context window. Existing scaffolds mitigate it wi…

HuggingFace Papers
论文 / 方法
6/21 20:00
Causal Discovery in the Era of Agents

Recent attempts to combine large language models (LLMs) with causal discovery ask models to infer pairwise directions, propose graph structures, or inject language-model outputs as priors and constrai…

HuggingFace Papers
论文 / 方法
6/21 20:00
Critique of Agent Model

What is an agent? What constitutes agency? With the rise of Large Language Model (LLM) systems marketed as ``coding agents'', ``AI co-scientists'', and other ``agentic" tools that promise to drive up…

HuggingFace Papers
设计 / 产品
6/21 19:52
anthony-chaudhary/fak

fak — the Fused Agent Kernel: one Go binary that turns a tool-using agent (Claude Code, Codex, Cursor, any OpenAI/Anthropic/MCP client) into a managed agent: ca…

GitHub
设计 / 产品
6/20 15:01
redevops-io/sidekick

Local coding-agent orchestrator — DAG of auto-approved, git-worktree-isolated sub-sessions across LLM providers (Claude/Kimi/Grok/DeepSeek/local). AGPL-3.0.

GitHub
Skill / 资源
6/19 22:45
Quoting Sean Lynch

The real valuable capability MCP offers over skills/CLI is isolating the auth flow outside of the agent’s context window, and potentially out of the harness completely. [...] Maybe…

AI 点评 · MCP将认证流程独立于智能体上下文,这种架构创新对提升安全性与扩展性意义重大。

Simon Willison
设计 / 产品
6/19 20:36
raiyanyahya/recall

Stop wasting tokens and re-explaining your project every session. Recall gives Claude Code durable memory — entirely offline.

GitHub
设计 / 产品
6/19 20:36
raiyanyahya/recall

Stop wasting tokens and re-explaining your project every session. Recall gives Claude Code durable memory — entirely offline.

GitHub
设计 / 产品
6/19 16:43
umacloud/umadev

UmaDev: A coding agent that works like a real dev team, commanding the Claude Code / Codex / OpenCode you already use.

GitHub
设计 / 产品
6/19 16:43
umacloud/umadev

UmaDev: A coding agent that works like a real dev team, commanding the Claude Code / Codex / OpenCode you already use.

GitHub
行业信号
6/19 15:59
强推 AI 引用户反感,谷歌 AI 建议用户不想看 AI 就用 DuckDuckGo

IT之家 6 月 19 日消息,谷歌目前正全力推进 AI 生态的建设,并在搜索引擎中强行加入 AI 智能体。 DuckDuckGo 官方今日在 X 上晒出了一张截图,显示谷歌 AI 概览正引导那些讨厌 AI 的用户前往 DuckDuckGo 的“No AI Search”页面,还提到了可调低 AI 体验强度的浏览器设置。 PiunikaWeb 测试发现,当用…

AI 点评 · 谷歌强推AI反遭打脸,AI自己建议用户用竞品,暴露了产品逻辑矛盾。

IT之家
设计 / 产品
6/19 15:37
印度首富安巴尼:印度必须成为 AI 的创造者和全球领导者

IT之家 6 月 19 日消息,印度首富、信实工业集团董事长穆克什 · 安巴尼希望,把公司打造为当地 AI 产业的代表,并把 AI 服务带入 电话、移动应用和智能家居 。 当地时间 19 日(今天),信实工业举行了年度股东大会,并发布 AI 通话助手 Jio Call Agent。Jio Call Agent 可以加入电话通话, 自动转录对话、生成摘要 ,还…

AI 点评 · 印度首富亲自押注,AI本土化野心显露,或重塑全球科技版图。

IT之家
设计 / 产品
6/19 13:24
sums001/Windows-Copilot-API

Reverse engineered Windows Copilot into an OpenAI-compatible API. Access GPT-4 and GPT-5 models through a simple REST interface without API keys or billing.

GitHub
设计 / 产品
6/19 13:24
sums001/Windows-Copilot-API

Reverse engineered Windows Copilot into an OpenAI-compatible API. Access GPT-4 and GPT-5 models through a simple REST interface without API keys or billing.

GitHub
设计 / 产品
6/19 11:58
Green-PT/honey-for-devs

Honey (I Shrunk the AI) by GreenPT: a cross-tool coding skill that cuts AI coding-agent token usage and LLM API costs — write less code, less prose, and denser…

GitHub
设计 / 产品
6/19 11:58
Green-PT/honey-for-devs

Honey (I Shrunk the AI) by GreenPT: a cross-tool coding skill that cuts AI coding-agent token usage and LLM API costs — write less code, less prose, and denser…

GitHub
设计 / 产品
6/19 11:12
adepeju4/attest

Evidence-grounded evaluation for AI agents — verifies each claim against the agent's real tool outputs (constrained, evidence-grounded model judgment, not holis…

GitHub
设计 / 产品
6/19 11:12
adepeju4/attest

Evidence-grounded evaluation for AI agents — verifies each claim against the agent's real tool outputs (constrained, evidence-grounded model judgment, not holis…

GitHub
设计 / 产品
6/19 06:34
Karovia/fullstack-ai-agent-roadmap

🎯 从零基础到 AI Agent 全栈工程师 · 110 个详细教程 · 58 万字 · 400+ GitHub 项目精选 · Obsidian 友好 · 中文

GitHub
行业信号
6/18 14:49
Launch HN: TesterArmy (YC P26) – Agents that test web and mobile apps

Hey HN - we’re Oskar, Szymon, and Piotr, and we’re building TesterArmy ( https://tester.army ). TesterArmy is an agentic testing platform that runs end-to-end checks before deploym…

AI 点评 · 用AI代理自动执行端到端测试,大幅提升应用发布前的质量保障效率。

Hacker News
设计 / 产品
6/18 06:18
shy3130/tickflow-stock-panel

自托管、零运维的 A 股「选股 + 监控 + 回测」量化工作台 | 基于 TickFlow 数据源 | LLM能力驱使策略定制+个股分析+复盘 | 自由接入第三方数据源与个性化扩展数据 | 个人开源 ,非TickFlow官方项目

GitHub
设计 / 产品
6/18 06:18
shy3130/tickflow-stock-panel

自托管、零运维的 A 股「选股 + 监控 + 回测」量化工作台 | 基于 TickFlow 数据源 | LLM能力驱使策略定制+个股分析+复盘 | 自由接入第三方数据源与个性化扩展数据 | 个人开源 ,非TickFlow官方项目

GitHub
行业信号
6/18 01:34
北大科学家下场做脑机接口,种子轮融了近亿元

文 | 孙小雯 访谈 / 编辑 | 海若镜 「暗涌Waves」独家获悉,侵入式脑机接口公司「芯生视界」近日完成近亿元人民币种子轮融资。本轮融资由经纬创投领投,星连资本、燕缘创投、水木创投跟投。 当下,侵入式脑机接口已经在治疗瘫痪、脑控外设等医疗场景落地,验证长期植入的安全、有效。与此同时,AI Agent和具身智能技术加速进化,也放大了市场对脑机接口的期待:…

36氪
Skill / 资源
6/17 20:35
Get back hours every day with autonomous agents in Amazon Quick

Today, Quick gets even more powerful: new autonomous agents that work continuously on your behalf, an activity feed that helps you prioritize your most important work, and the abil…

AI 点评 · 亚马逊Quick新增自主代理,能持续替用户工作,大幅提升效率,值得关注。

AWS ML
设计 / 产品
6/17 14:31
LING71671/open-reverselab

Open-source reverse engineering lab: 197-article knowledge base + MCP tools + CTF/APK/PE automation toolchain. Agent-native. Note:由于场景原因,目前有让几乎所有(除fable5)AI都会越…

GitHub
设计 / 产品
6/17 13:01
xorbitsai/xrouter-llm

A prompt-aware LLM router that predicts which models can complete each request, then selects the cheapest capable one: 53.2% lower cost and +1.9 pts completion…

GitHub
论文 / 方法
6/16 20:00
Playful Agentic Robot Learning

Current agentic robot systems can write executable Code-as-Policy programs, observe feedback, and revise behavior across multiple attempts, but they remain largely task-driven: reusable skills are acq…

HuggingFace Papers
设计 / 产品
6/16 09:13
Alisa0808/vibe-creating-skill

Open-source, bilingual AI video-prompt skill — rewrite ideas into model-ready text-to-video prompts. A portable Agent Skill (Claude Code, Codex, OpenClaw, Herme…

GitHub
设计 / 产品
6/16 06:18
yzhao062/auditable

Audit any agent decision across its past, present, and future, on one typed graph.

GitHub
设计 / 产品
6/16 06:18
yzhao062/auditable

Audit any agent decision across its past, present, and future, on one typed graph.

GitHub
设计 / 产品
6/15 23:07
macOS 26.4 为何拦截部分终端命令?苹果解释背后安全触发机制

IT之家 6 月 16 日消息,苹果公司昨日(6 月 15 日)更新支持文档,解释称在 macOS 26.4 系统中, 若用户不常用终端(Terminal),且命令来自网站、聊天智能体、消息或邮件应用,系统可能阻止粘贴。 科技媒体 9to5Mac 指出,用户此前终端里粘贴命令后,系统会先给出安全警告,提示内容可能含有恶意代码。很多人只知道有这个拦截机制,但不…

AI 点评 · 安全机制升级针对非高频操作,防范恶意代码粘贴执行,体现系统防护精细化。

IT之家
设计 / 产品
6/15 22:56
古尔曼:苹果有望推出 AI 智能体,让 Siri 自主操作 iPhone 和 Mac 软件

IT之家 6 月 16 日消息,彭博社记者马克 · 古尔曼认为,苹果最终可能推出一款产品,直接对标 OpenClaw—— 这是一套智能体 AI 系统,能够代用户自主操作各类软件。 古尔曼在其专栏《Power On》中撰文表示,他预计苹果会研发一套系统,可全权代表用户操作 iPhone、iPad 与 Mac 端的各类软件。这一预测的依据,是苹果 Siri 工程…

AI 点评 · Siri从语音助手进化为自主操作系统的AI代理,标志苹果在智能体赛道的关键布局。

IT之家
论文 / 方法
6/15 20:00
CEO-Bench: Can Agents Play the Long Game?

Language model agents are becoming proficient executors at isolated, short-horizon tasks such as software engineering and customer service. Yet real-world challenges require a combination of sophistic…

HuggingFace Papers
论文 / 方法
6/15 17:59
Context-Aware RL for Agentic and Multimodal LLMs

Large language models (LLMs) often fail when answering requires identifying a small but decisive piece of evidence within a long or complex context, such as a single line in a tool trace or a subtle d…

arXiv
Skill / 资源
6/15 17:19
datasette-agent 0.3a0

Release: datasette-agent 0.3a0 New tool, execute_write_sql , which requests user approval and then writes to a database - taking user permissions into account. #27 I added a mechan…

Simon Willison
设计 / 产品
6/15 09:15
volcengine/ark-cli

The fastest way to put Volcengine Ark in your terminal and your AI agent — go from prompt to generated media, multimodal answer, or deployed endpoint in a sin…

GitHub
设计 / 产品
6/14 23:02
Egoist-Machines/etchplan

Compile an AI agent's repeated workflows into deterministic, auditable routines that replay for free, with a fallback to the agent.

GitHub
论文 / 方法
6/14 20:00
ProCUA-SFT Technical Report

Training computer-use agents (CUAs) -- models that interact with graphical desktops through screenshots and keyboard/mouse actions -- requires large-scale, diverse trajectory data collected in full de…

HuggingFace Papers
设计 / 产品
6/14 00:36
001TMF/harness-forge

Turn Claude Code into its own Meta-Harness — a skill that evolves the scaffolding around a fixed model (memory, retrieval, context, prompts) via a native propos…

GitHub
设计 / 产品
6/13 23:28
谷歌推出搜索智能体功能,可主动帮你盯全网信息

IT之家 6 月 14 日消息,日常上网搜索时,往往需要停下手中的事、打开标签页,主动去查找最新信息。如今,谷歌打算彻底改变这种使用模式。继在 2026 年谷歌开发者大会上首次预告后,谷歌现已正式在 AI 模式中推出搜索智能体功能。 此次升级将传统搜索引擎转变为可在后台静默运行的主动式助手。 首批上线的是信息智能体功能,它会主动全网监测信息,无需用户手动检索…

AI 点评 · 从被动检索到主动监测,AI搜索正从工具进化为管家,颠覆传统上网模式。

IT之家
设计 / 产品
6/13 22:09
UCSC-VLAA/VisualClaw

Official Implementation of VisualClaw: A Real-Time, Personalized Agent for the Physical World

GitHub
设计 / 产品
6/13 14:58
知名会计师事务所毕马威 AI 行业报告被指是“AI 写的”:充斥幻觉、错误百出

IT之家 6 月 13 日消息,去年 10 月,毕马威曾发布《总体体验:在智能体 AI 时代重新定义卓越》报告,讨论企业如何利用 AI 满足客户需求。然而据英国《金融时报》12 日报道,这份报告后来被“抓包”充斥 AI 幻觉:报告列举的多个智能体 AI 案例, 要么并不存在,要么并不具备毕马威所描述的能力 。 AI 内容检测工具开发商 GPTZero 的调查…

AI 点评 · 专业机构因依赖AI工具反被AI误导,暴露了当前生成式模型在事实核查上的致命短板。

IT之家
模型 / Agent
6/13 14:56
智谱 AI 编程工具 ZCode 3.0 版本发布:切换自研 ZCode Agent 内核,深度适配 GLM-5.2

IT之家 6 月 13 日消息,智谱今日发布了 AI 编程工具 ZCode 3.0 新版本, 深度适配 GLM-5.2 。 官方表示,ZCode 3.0 全面 切换自研 ZCode Agent 内核 。针对满血 GLM 深度优化长程推理、工具调用和大型工程执行链路,整体任务完成效果已显著优于第三方 Agent; 后续版本将聚焦自研 Agent 体验,不再内置…

AI 点评 · 自研内核替代第三方方案,长程推理与工程执行能力显著提升。

IT之家
设计 / 产品
6/13 14:22
dzcmemory-web/bazi-ziwei-skill

AI 八字 + 紫微斗数排盘与综合印证 Skill:算法精准排盘(不靠 LLM 猜),三种分析模式,一键生成水墨风 HTML 命盘海报。兼容 Claude / Codex / Cursor / Workbuddy 等 SKILL.md Agent。

GitHub
行业信号
6/13 09:44
Show HN: Paca – Lightweight Jira alternative for human-AI collaboration

I built Paca out of pure passion—a free and lightweight Jira alternative written in Go where humans and AI agents work together as equal teammates to plan sprints and assign tasks…

AI 点评 · 轻量级AI协作工具挑战Jira,用Go开发的开源项目实现人机平等分工,值得关注。

Hacker News
设计 / 产品
6/13 03:15
eli-labz/Third-Eye

A production-grade OSINT platform that provides situational awareness across multiple intelligence domains.

GitHub
模型 / Agent
6/12 21:00
NVIDIA Blackwell Leads on First Agentic AI Infrastructure Benchmark

AgentPerf from Artificial Analysis, the industry’s first agentic AI benchmark, gives developers, enterprises and infrastructure providers a clear way to compare systems for agentic…

AI 点评 · NVIDIA在首个智能体AI基准测试中夺冠,为行业选择基础设施提供关键参考。

NVIDIA
设计 / 产品
6/12 15:19
华为余承东:鸿蒙 HarmonyOS 成为中国第二大智能手机操作系统

IT之家 6 月 12 日消息,在今日的华为开发者大会 HDC 2026 上,华为常务董事、产品投资评审委员会主任、终端 BG 董事长余承东发布了新一代鸿蒙 HarmonyOS 7 操作系统,围绕互联、智能、安全、流畅、空间感五个维度进行升级,同时迈向智能体时代。 余承东在 HDC 2026 现场宣布, 鸿蒙 HarmonyOS 已成为中国第二大智能手机操作…

AI 点评 · 标志着国产操作系统突破安卓和iOS垄断,生态建设进入新里程碑。

IT之家
设计 / 产品
6/12 00:52
DietrichGebert/ponytail

Makes your AI agent think like the laziest senior dev in the room. The best code is the code you never wrote.

GitHub
行业信号
6/11 23:24
OpenAI 收购初创公司 Ona,强化编程助手 Codex

IT之家 6 月 12 日消息,OpenAI 昨天宣布收购初创公司 Ona,为编程助手 Codex 提供安全、预配置云环境。 IT之家从官方新闻稿获悉,Ona 的技术将帮助 Codex 执行持续时间更长的任务,并帮助用户将 AI 智能体部署到生产环境。 同时, Ona 的全新技术将帮助企业更好掌控基础设施 、 数据资产和安全边界 ,让 Codex 能够在安全…

AI 点评 · 收购Ona补齐Codex短板,从代码生成迈向安全部署,标志AI编程助手进入企业级实战阶段。

IT之家
论文 / 方法
6/11 20:00
LLM Agents Can See Code Repositories

Coding agents powered by large language models have demonstrated strong performance on software engineering tasks. Yet most agents consume repositories almost entirely as text, which differs from how…

HuggingFace Papers
论文 / 方法
6/11 17:58
Agents-K1: Towards Agent-native Knowledge Orchestration

Current LLM-based research agents have advanced through agent orchestration, yet largely overlook scientific knowledge orchestration. Existing works often reduce papers to abstracts, surface mentions,…

arXiv
论文 / 方法
6/11 17:47
Recursive Agent Harnesses

Recursive language models (RLMs) showed that recursion over model calls is an effective strategy for long-context reasoning, and production coding agents have begun to write code that spawns subagents…

arXiv
Skill / 资源
6/11 15:49
Evaluate AI agents systematically with Agent-EvalKit

Agent-EvalKit is an open-source toolkit (Apache 2.0) that makes this evaluation infrastructure available by integrating with AI coding assistants, including Claude Code, Kiro CLI,…

AI 点评 · 开源工具填补AI智能体评估基础设施空白,让开发者能系统化测试性能。

AWS ML
设计 / 产品
6/11 12:18
omnigent-ai/omnigent

Omnigent is an open-source AI agent framework and meta-harness: orchestrate Claude Code, Codex, Cursor, Pi, and custom agents — swap harnesses without rewriting…

GitHub
设计 / 产品
6/11 09:00
HarryHsing/OmniAgent

OmniAgent (ICML 2026): the first native omni-modal agent for active video perception — a 7B agent that beats Qwen2.5-VL-72B with 73% fewer frames on LVBench.

GitHub
模型 / Agent
6/11 00:00
OpenAI to acquire Ona

OpenAI plans to acquire Ona to expand Codex with secure, persistent cloud environments, enabling long-running AI agents across enterprise workflows.

OpenAI
Skill / 资源
6/10 23:57
datasette-agent 0.2a0

Release: datasette-agent 0.2a0 Highlights from the release notes: Tools can now ask the user questions mid-execution. Tools that declare a context parameter receive a ToolContext o…

Simon Willison
论文 / 方法
6/10 20:00
MiniMax Sparse Attention

Ultra-long-context capability is becoming indispensable for frontier LLMs: agentic workflows, repository-scale code reasoning, and persistent memory all require the model to jointly attend over hundre…

HuggingFace Papers
论文 / 方法
6/10 17:47
APPO: Agentic Procedural Policy Optimization

Recent advances in agentic Reinforcement Learning (RL) have substantially improved the multi-turn tool-use capabilities of large language model agents. However, most existing methods assign credit ove…

arXiv
设计 / 产品
6/10 16:30
DizzyMii/fable-skills

Six Claude Code skills that harden Opus 4.8 toward frontier behavior — written by Fable 5, pressure-tested on the target model with transcripts included.

GitHub
设计 / 产品
6/10 07:46
chuspeeism/dashiAI-ppt-skill

An AI-agent skill that generates browser-editable presentations from multiple visual themes, exportable to HTML, PDF, and PPTX.

GitHub
行业信号
6/9 23:41
高知特等六大集成商联手推广Rubrik AI安全平台

Rubrik公司周二在Rubrik FORWARD大会上宣布,六家全球系统集成商将为企业客户部署其专为Anthropic Claude Code打造的Rubrik Agent Cloud平台。加入“沙漏计划”的合作伙伴包括高知特、德勤、LTM、HCL科技、NTT Data以及威普罗,这些集成商将把该平台整合至各自的网络安全与数字化转型服务体系中。(新浪财经)

AI 点评 · 六大集成商联合推广,标志AI安全平台从技术验证进入规模化落地阶段。

36氪
行业信号
6/9 23:35
毕马威与微软扩大合作,向逾27万员工全面铺开Copilot并引入Agent 365治理框架

毕马威与微软6月9日宣布扩展全球战略合作关系,聚焦企业级AI智能体的规模化部署。根据协议,微软365 Copilot将向毕马威全球逾27.6万名专业人员全面推广;与此同时,毕马威将采用微软Agent 365平台,对其全球组织内及客户端的AI智能体实施统一管理、监控与安全治理。(界面)

AI 点评 · 毕马威27万员工全面接入Copilot,并首创Agent治理框架,预示企业级AI从工具应用迈入系统化

36氪
设计 / 产品
6/9 22:07
PolyHelper/polyhelper

Self-evolving cognitive AI exoskeleton. 10+ frontier models, 245 consensus methods, governed autonomous agents. Automotive, medical, legal, accessibility. 9.3M…

GitHub
Skill / 资源
6/9 21:35
Setting a custom price for a model in AgentsView

TIL: Setting a custom price for a model in AgentsView I've been really enjoying AgentsView by Wes McKinney as a tool for exploring my token usage across different coding agents run…

AI 点评 · 自定义AI模型定价功能,让用户按需付费,提升灵活性与成本控制。

Simon Willison
行业信号
6/9 12:48
氪星晚报 |腾讯、阿里等入股脑机接口研发商阶梯医疗;飞猪:端午假期入境游预订量同比增长超6倍;1—5月全国期货市场累计成交额同比增长40.13%

大公司: 猫眼娱乐:作为首批内测开发者接入微信AI生态布局 36氪获悉,猫眼娱乐宣布作为微信AI生态首批内测开发者之一,旗下小程序接入微信AI Agent生态,借助微信AI Agent能力,为用户提供影片演出推荐、附近影院筛选、智能选座、一键支付等服务。 京东方A:控股子公司拟终止向不特定合格投资者公开发行股票并撤回申请文件 36氪获悉,京东方A公告,公司控…

36氪
设计 / 产品
6/9 11:20
wanshuiyin/ARIS-Movie-Director

Agentic, long-horizon visual generation: a fuzzy story → a cross-model-audited image-based movie. Brings ARIS's research-wiki + multi-agent debate to multimodal…

GitHub
设计 / 产品
6/9 06:28
cobusgreyling/loop-engineering

Practical patterns, starters & CLI tools for loop engineering with AI coding agents. Design systems that prompt and orchestrate agents (inspired by Addy Osmani…

GitHub
论文 / 方法
6/8 20:00
Kwai Keye-VL-2.0 Technical Report

We introduce Kwai Keye-VL-2.0-30B-A3B, an open-source Mixture-of-Experts (MoE) multimodal foundation model designed to advance long-video understanding and agentic intelligence. To address the challen…

HuggingFace Papers
论文 / 方法
6/8 17:53
FASE: Fast Adaptive Semantic Entropy for Code Quality

Multi-agent code generation offers a promising paradigm for autonomous software development by simulating the human software engineering lifecycle. However, system reliability remains hindered by LLM…

AI 点评 · 提出自适应语义熵指标,精准衡量代码质量,突破多智能体编程可靠性瓶颈。

arXiv
论文 / 方法
6/8 17:35
SIGA: Self-Evolving Coding-Agent Adapters for Scientific Simulation

Advanced scientific simulators expose specialized input languages that turn simulation goals into executable configurations, but learning them can cost domain scientists hours to days. We study simula…

AI 点评 · 用代码生成适配器降低科研模拟门槛,让科学家专注研究而非编程。

arXiv
论文 / 方法
6/8 17:27
iOSWorld: A Benchmark for Personally Intelligent Phone Agents

A useful phone agent needs to be personally intelligent. It should reason over a user's identity, history, and preferences as they exist on the device, not just follow isolated instructions in an impe…

AI 点评 · 首个衡量手机智能体个性化推理能力的基准,填补了当前AI助手忽略用户身份与历史数据的评测空白。

arXiv
设计 / 产品
6/8 09:01
腾讯汤道生评价姚顺雨、混元 3和元宝

文|王毓婵 编辑|张雨忻 6月5日,腾讯云AI产业应用大会最受外界关注的是什么? 毫无疑问,是汤道生与姚顺雨的对话。 在这场发布了一系列覆盖20多个垂直场景Agent的大会上,因为产品过于To B,也没有提及大家最关注的“微信AI”,导致外界的关注重心几乎全部被那场对话吸引走。 腾讯集团高级执行副总裁、云与智慧产业事业群CEO汤道生,在对谈中问腾讯首席AI科…

36氪
设计 / 产品
6/8 08:42
OtterMind/Nubase

🔥🔥🔥 Turn AI-written code into real apps. Nubase is an open-source, AI-native backend platform for AI Coding, agentic applications, and modern product teams:…

GitHub
模型 / Agent
6/8 08:10
tigicion/dao-code

Open-source TypeScript terminal coding agent for DeepSeek-V4 — builds on DeepSeek's strong price-performance and ultra-cheap cache pricing, engineering byte-sta…

GitHub
设计 / 产品
6/8 04:20
蚂蚁集团推出海外AI支付解决方案

36氪获悉,近日,蚂蚁国际面向全球电子钱包、超级应用和数字银行等移动支付服务正式推出移动智能体协议(Agentic Mobile Protocol,简称AMP),解决全球消费者在AI智能体(agents)中购物时支付便捷、安全、有保障,以及商家跨市场互联互通等智能体商业全球化运营的关键难题。

AI 点评 · 蚂蚁国际首创移动智能体协议,破解AI购物支付与全球互联难题,引领跨境支付新标准。

36氪
Skill / 资源
6/7 23:56
datasette-agent-edit 0.1a0

Release: datasette-agent-edit 0.1a0 I'm planning several plugins for Datasette Agent which can make edits to existing pieces of text - things like collaborative Markdown editing, u…

AI 点评 · AI文本编辑插件迈出关键一步,轻松实现多人协作修改,值得关注。

Simon Willison
行业信号
6/7 23:03
京东与腾讯就AI Agent将达成重要合作

6月7日晚间消息,京东与腾讯将围绕AI Agent展开深度合作。依托京东的商品供应链、履约服务能力及腾讯的生态入口优势,双方将共同打造跨场景的智能化服务新范式,推动AI Agent从单点应用走向生态协同。(新浪科技)

AI 点评 · 京东与腾讯联手,AI Agent从单点走向生态协同,或重塑电商与社交的智能服务边界。

36氪
设计 / 产品
6/7 19:58
fkiene/llmtrim

Local proxy that compresses your LLM API requests so you pay less, with no change to the answers. Trims wasted tokens from prompts, history, tool output, and co…

GitHub
模型 / Agent
6/7 14:37
支持调用 Deepseek、Kimi 等模型,Agentic 华为云入口“智果园”发布

IT之家 6 月 7 日消息,华为云现已针对 Agentic AI 时代发布全新云入口“智果园”,新产品支持云码道 CodeArts 代码智能体、华为云 OfficeAce 办公智能体和 WorkAgent 文档智能体。 据介绍,智果园拥有开发、办公等多种关键行业的智能体,可通过智果 AgentArts 平台打造更加实用的智能体,并通过 Skills、AI…

AI 点评 · 华为云整合多模型,降低企业AI应用门槛,凸显生态开放与行业落地潜力。

IT之家
行业信号
6/7 10:23
消息称京东、腾讯联手,将围绕 AI Agent 展开合作

IT之家 6 月 7 日消息,据钛媒体今日消息, 京东与腾讯已于近期联手,将围绕 AI Agent 展开合作 。京东的商品供应链与履约服务体系,将与腾讯的入口资源进行对接。 此外,消息称 京东 AI Agent 与华为、OPPO、荣耀等多家主流终端厂商已进行对接 。通过 A2A(Agent to Agent)合作,用户可直接在各终端原生智能体的京东 AI A…

AI 点评 · 巨头联手布局AI Agent,消费场景与社交入口的深度融合将加速智能体商业化落地。

IT之家
模型 / Agent
6/7 06:24
patraxo/ltx2-vidgen-skill

Own your AI video pipeline. LTX-2.3 (22B) self-hosted on your Modal GPU via a Claude Code skill — t2v, i2v, keyframes, v2v + synced audio. ~$0.02 per 5s clip, i…

GitHub
设计 / 产品
6/7 04:23
Light0305/Light-skills

Light — 全流程科研技能包:28 个技能覆盖文献调研到投稿全流程,配套 9 个可核查知识库。适配主流 AI 编程客户端。

GitHub
设计 / 产品
6/7 04:23
Light0305/Light-skills

An AI workflow skill pack for research, competitions, and innovation projects.

GitHub
设计 / 产品
6/6 13:49
微软想让用户对自家智能体 Scout 上瘾?CEO 纳德拉否认

IT之家 6 月 6 日消息,微软在 Build 2026 上与高通联合发布 Project Solara。Project Solara 主打“智能体优先计算”,系统内部运行 Agent Shell,并能动态加载、调整多个基于云端的 AI 智能体。 微软 CEO 萨提亚 · 纳德拉表示:AI 智能体已经 不再只是普通 AI 助手 。“真正的平台转变正在发生。…

AI 点评 · 微软新项目争议背后,揭示AI从“助手”到“智能体”的平台级变革。

IT之家
论文 / 方法
6/5 17:59
Agentopia: Long-Term Life Simulation and Learning in Agent Societies

Humans learn from social life. Simulating this process with LLM-powered agents represents a promising research direction, raising a natural question: whether LLMs can learn from such simulated social…

AI 点评 · AI社会模拟研究突破:用LLM代理模拟人类长期社交学习,探索机器从社会互动中进化。

arXiv
论文 / 方法
6/5 17:59
MemDreamer: Decoupling Perception and Reasoning for Long Video Understanding via Hierarchical Graph Memory and Agentic Retrieval Mechanism

Current Vision-Language Models struggle with hours-long videos because processing full-length visual sequences induces prohibitive token explosion and attention dilution. To overcome this, we introduc…

AI 点评 · 分层记忆架构破解长视频理解瓶颈,用图记忆与智能检索分离感知推理,显著降低计算成本。

arXiv
论文 / 方法
6/5 17:51
Accelerated Decentralized Stochastic Gradient Descent for Strongly Convex Optimization

Decentralized stochastic optimization is a fundamental paradigm for large-scale learning over networks, where agents communicate only with their neighbors and no central coordinator is required. For s…

AI 点评 · 加速去中心化随机梯度下降,破解大规模网络学习效率瓶颈,强凸优化领域关键突破。

arXiv
论文 / 方法
6/5 17:45
How AI Agents Reshape Knowledge Work: Autonomy, Efficiency, and Scope

Frontier AI systems are bridging the gap between intelligence and utility by shifting from conversational assistants to autonomous agents that execute tasks end to end. Using production data from Perp…

AI 点评 · AI代理将知识工作从辅助变为自主执行,效率与范围迎来质变。

arXiv
设计 / 产品
6/5 14:39
lamenting-hawthorn/SkillLoop

Standalone self-improvement harness for agent traces, memory, skills, evaluation, and fine-tuning exports

GitHub
设计 / 产品
6/5 14:32
paxlabs-inc/matrix-core

Matrix is the cognition and UX layer on top of Paxeer Network. It turns natural-language requests from non-developers into a typed, inspectable, correctable Int…

GitHub
行业信号
6/5 09:00
The Meta hack shows there’s more to AI security than Mythos

On June 5, 404 Media reported that attackers had been using Meta’s AI customer support agent to steal Instagram accounts. Their approach was simple: They asked the agent to link th…

AI 点评 · Meta AI客服漏洞暴露了安全盲区,提醒行业警惕简单攻击手法。

MIT Tech Review
设计 / 产品
6/5 07:57
Illuminfti/flow

A dynamic-workflow engine for any agent, any model: concurrent leaves, cost-aware routing, schema enforcement, crash-resume, bounded iterate-to-goal loops with…

GitHub
设计 / 产品
6/5 07:56
hoolulu/deep-research

深度调研报告生成 Skill — 一条命令,十分钟出券商级深度调研报告 / Professional deep research report generation Skill · Supports 19 languages

GitHub
设计 / 产品
6/5 03:39
微信AI对手机厂商打开一道窄门|焦点分析

文|王毓婵 梁键强 编辑|张雨忻 昨日,腾讯客服回应称,微信正在与华为、小米、荣耀、OPPO、vivo等手机厂商合作推出A2A助手能力,目前已有多家厂商完成接入。 “您可以通过对应手机系统的AI助手发起 微信音视频通话 或 向指定好友发送消息 。该功能基于A2A(Agent-to-Agent)协作机制,数据安全与隐私通过双重授权机制保障。合作旨在将微信高频沟…

AI 点评 · 微信与手机厂商打通AI协作,标志社交场景向多终端智能体互联迈出关键一步。

36氪
设计 / 产品
6/4 23:46
8点1氪丨分析师曝苹果Vision Pro产品线被移除;黄仁勋将首次亮相综艺节目;粉笔CEO就骂人大学生“活该找不到工作”道歉

今日热点导览 微信正与手机厂商合作推出Agent-to-Agent助手能力 擅用“LABUBU”相近标识商业推广,泡泡玛特告奈雪的茶获赔32万 腾讯客服回应与华为、小米等合作 神农旅游集团就国道收费进行道歉 苹果智能眼镜推迟至2029年,无显示屏AI眼镜仍将于2027年推出 TOP 3 大新闻 分析师曝苹果Vision Pro产品线被移除,2027年推AI眼…

36氪
设计 / 产品
6/4 23:31
苹果批准首个 iMessage AI 智能体,Poke 可回邮件也能设提醒

IT之家 6 月 5 日消息,科技媒体 Appleinsider 昨日(6 月 4 日)发布博文, 报道称苹果批准 Poke 成为首个接入 Apple Messages for Business(苹果商务消息)平台的第三方 AI 智能体。 Apple Messages for Business 原本服务企业客服沟通的通道,苹果调整后开始承载更主动的 AI 助…

AI 点评 · 苹果开放iMessage接入AI,标志其生态向第三方智能体迈出关键一步。

IT之家
论文 / 方法
6/4 20:00
The Cold-Start Safety Gap in LLM Agents

Are tool-calling LLM agents equally safe throughout a conversation? We discover they are not: agents are most vulnerable at the very start of a session and become substantially safer after a few regul…

HuggingFace Papers
模型 / Agent
6/4 16:59
NVIDIA Nemotron 3 Ultra now available on Amazon SageMaker JumpStart

Deploy NVIDIA Nemotron 3 Ultra on Amazon SageMaker JumpStart. Get 5x faster inference and 30% lower cost for agentic AI workloads with this frontier reasoning model.

AI 点评 · 英伟达顶级推理模型登陆云平台,5倍推理加速与30%成本降低,企业AI部署门槛再降。

AWS ML
设计 / 产品
6/4 09:37
氪星晚报 |OpenAI首席财务官谈公司AI设备:今年年底前将正式发布;腾讯客服回应与华为、小米等合作;香港推出首个生产力级超级智能体

大公司: OpenAI首席财务官谈公司AI设备:今年年底前将正式发布 OpenAI首席财务官Sarah Friar日前在受访时透露,已经亲自体验过OpenAI的AI设备。她表示到“今年年底之前”OpenAI将正式发布这款产品。此前,OpenAI曾在一份文件中表示,预计最早也要到2027年2月才会开始发货。(财联社) LG Innotek计划扩建半导体基板工厂…

36氪
设计 / 产品
6/3 23:31
IT早报 0604:豆包将推专业版,基础功能免费;估值 1.77 万亿美元,SpaceX 发行价敲定;腾讯人士称微信智能体上线时间暂未定;大疆否认 Pocket 4 饥饿营销...

“IT早报”时间,大家好,现在是 2026 年 6 月 4 日星期四,今天的重要科技资讯有: 1、豆包:计划针对专业人群生产力需求推出豆包专业版,基础功能保持免费 豆包声明称,对于广大用户日常使用的豆包功能,包含搜索问答、写作生图、以及语音和视频对话等,将保持目前的免费服务,保证用户使用体验和习惯不受影响。>> 查看详情 2、SpaceX 敲定 IPO 发行…

AI 点评 · 多条科技资讯集中释放,豆包专业版与免费策略、SpaceX估值与上市动态均值得关注。

IT之家
设计 / 产品
6/3 18:55
tastyeffectco/sandboxes

Self-hosted dev sandboxes with preview URLs. One command. No Kubernetes, perfect for coding agents and Saas factories

GitHub
设计 / 产品
6/3 18:55
tastyeffectco/sandboxd

Self-hosted dev sandboxes with preview URLs. One command. No Kubernetes, perfect for coding agents and Saas factories

GitHub
行业信号
6/3 17:45
As AI gets better, it reveals an empty promise

This week we've got tandem hands-ons with Google's new Gemini AI agent - Spark - from my colleagues David Pierce and Jay Peters. Their takeaways are similar: It's so effective that…

AI 点评 · AI能力提升反暴露技术天花板的矛盾,揭示行业深层困境。

The Verge
设计 / 产品
6/3 14:56
eddyzzl/marvis-risk-agent

MARVIS-Agent: all-purpose credit risk agent for model development, validation, data processing, feature engineering, and strategy workflows.

GitHub
设计 / 产品
6/3 08:07
davanstrien/uv-scripts-for-ai

Self-contained UV scripts for data & ML tasks — OCR, vision, audio & more — run one in a command, locally or on Hugging Face Jobs. Built for humans and agents.

GitHub
设计 / 产品
6/3 04:26
Nigh/show-me-the-story

Self-hosted AI novel generator: single Go binary + web UI. OpenAI-compatible API → outline → chapter-by-chapter writing with review, foreshadowing, fact-check,…

GitHub
设计 / 产品
6/3 02:34
腾讯AI产业应用大会在即,即将发布系列智能体应用新品

36氪获悉,据腾讯云官号账号信息,腾讯2026AI产业应用大会即将在北京举办。作为腾讯年度最重要的AI产品发布平台,此次将发布系列智能体应用新品,并将公布infra等基础设施升级新进展。与此同时,腾讯集团高级执行副总裁、云与智慧产业事业群CEO汤道生将于腾讯AI首席科学家姚顺雨同台对话,解读AI下半场腾讯在AI赛道的最新布局和思考。

AI 点评 · 腾讯年度AI战略窗口,智能体新品与基础设施升级同步亮相,产业布局信号明确。

36氪
设计 / 产品
6/3 02:20
腾讯人士:目前无法确定微信 AI 智能体何时推出

IT之家 6 月 3 日消息,据财经杂志报道,腾讯人士表示,目前无法确定微信 AI 智能体何时推出,其上线时间很大程度上取决于监管方对智能体的审批进度,微信 14 亿的用户体量,合规流程可能比其他产品更加严格。关于微信智能体,腾讯相关负责人表示暂无回应。 IT之家注意到,此前英国《金融时报》报道称,微信将推出一款 AI(人工智能)智能体,计划最快将于本月启动…

AI 点评 · 监管审批成关键变量,14亿用户规模下的合规挑战值得关注。

IT之家
设计 / 产品
6/3 02:15
微信将推出一款AI智能体?腾讯人士回应

昨日,媒体报道称,微信将推出一款AI(人工智能)智能体,计划最快将于本月启动公开上线前所需的合规审批流程。腾讯人士表示,目前无法确定微信AI智能体何时推出,其上线时间很大程度上取决于监管方对智能体的审批进度,微信14亿的用户体量,合规流程可能比其他产品更加严格。关于微信智能体,腾讯相关负责人表示暂无回应。(财经)

AI 点评 · 微信14亿用户体量下,AI智能体合规审批进度成焦点,决定产品上线时间。

36氪
设计 / 产品
6/3 01:50
微软定调 Win11:打造成 AI 应用和智能体开发平台

IT之家 6 月 3 日消息,科技媒体 Windows Latest 今天(6 月 3 日)发布博文,报道称在 2026 年 Build 开发者大会上,微软明确 Windows 11 系统定位: 不再只是带 AI 功能的桌面系统,而是要成为 AI 应用和智能体的开发平台。 微软新方向涵盖智能体 Runtime、本地模型、Windows 原生 AI 接口、Li…

AI 点评 · 微软从用户工具转向开发者平台,AI生态野心浮出水面。

IT之家
设计 / 产品
6/2 23:14
郭明錤:黄仁勋高喊“重新发明 PC”口号凝聚市场共识,英伟达 RTX Spark 勾勒端侧 AI 智能体蓝图

IT之家 6 月 3 日消息,天风国际证券分析师郭明錤今天(6 月 3 日)在 X 平台发布推文,再次评论英伟达的 RTX Spark, 认为该处理器在未来 2 年内仍是小众产品,苹果在 WWDC 上对于设备端 AI 智能体的回应将是除 Siri 之外的另一个观察重点。 郭明錤表示英伟达 RTX Spark 处理器的核心看点不仅在于芯片本身, 更重要的是黄仁…

AI 点评 · 分析师视角揭示英伟达布局端侧AI的战略意图,市场影响值得关注。

IT之家
论文 / 方法
6/2 20:00
Personal AI Agent for Camera Roll VQA

We study the personal camera roll visual question answering setting. In this setting, a conversational AI assistant can access a user's personal camera roll and retrieve relevant photos to answer quer…

HuggingFace Papers
论文 / 方法
6/2 20:00
Agents' Last Exam

Recent AI systems have achieved strong results on a wide range of benchmarks, yet these gains have not translated into economically meaningful deployment across many professional domains. We argue tha…

HuggingFace Papers
模型 / Agent
6/2 19:28
datasette-agent-micropython 0.1a0

Release: datasette-agent-micropython 0.1a0 I want Datasette Agent to be able to generate and execute Python code safely. This alpha is looking promising so far. GPT-5.5 has so far…

AI 点评 · 结合AI代理与MicroPython,实现安全代码生成执行,为数据探索带来新可能。

Simon Willison
Skill / 资源
6/2 19:20
micropython-wasm 0.1a1

Release: micropython-wasm 0.1a1 Fixes for some limitations that emerged while I was trying to use this to build datasette-agent-micropython . Tags: python , sandboxing , webassembl…

AI 点评 · 在浏览器中运行MicroPython,为Python沙箱执行和Web应用开辟新可能。

Simon Willison
设计 / 产品
6/2 19:13
superloglabs/superlog

Open-source observability tool that uses AI agents to self-heal your software

GitHub
行业信号
6/2 18:19
Microsoft announces Scout, an autonomous AI agent built on OpenClaw

https://www.microsoft.com/en-us/microsoft-365/blog/2026/06/0... https://www.404media.co/microsoft-wants-to-make-people-addic... https://www.wired.com/story/meet-microsoft-scout-you…

AI 点评 · 微软自研AI代理Scout基于OpenClaw,标志着巨头在自主智能体领域的战略布局。

Hacker News
行业信号
6/2 18:00
Microsoft offers devs a better way to control AI agent behavior

The specification lets developer, compliance, and security teams define their own policies for agents to follow in portable policy files.

AI 点评 · 微软推出便携式策略文件,让开发者自主定义AI代理行为规范,提升安全可控性。

TechCrunch
行业信号
6/2 17:31
Microsoft’s Project Solara is an OS for AI agent gadgets

Microsoft just announced "Project Solara," a new OS designed for gadgets that run AI agents, at Build 2026. The company is calling it "a new platform built from the ground up to po…

AI 点评 · 微软专为AI智能体硬件打造操作系统,标志从软件到硬件生态的关键一步。

The Verge
设计 / 产品
6/2 12:48
pfwjrfp5hh-byte/WorkMesh

Open-source AI-era employment platform connecting skills, jobs, enterprises, governance, and AI agents.

GitHub
设计 / 产品
6/2 12:48
yangyunice/WorkMesh

Open-source AI-era employment platform connecting skills, jobs, enterprises, governance, and AI agents.

GitHub
行业信号
6/2 12:27
原华为盘古“90 后少帅”王云鹤离职创业,新公司“基元律动”获 1 亿美元估值融资

IT之家 6 月 2 日消息,据新浪科技今日报道,曾在华为主导盘古大模型研发的“90 后少帅”王云鹤,已于近期投身 AI Agent 领域创业,其新成立的公司“基元律动”已完成一轮估值达 1 亿美元的新融资。 王云鹤在今年 3 月末正式告别了工作近 9 年的华为。离职前,他最后的职务为华为诺亚方舟实验室主任、盘古大模型负责人,曾被誉为“盘古大模型少帅”和“天…

AI 点评 · 顶尖技术人才创业动向,折射AI Agent赛道资本热度与行业新趋势。

IT之家
设计 / 产品
6/2 12:03
英伟达 Spectrum- X 以太网硅光技术已全面量产,较传统网络能效提升 5 倍

IT之家 6 月 2 日消息,英伟达于 5 月 31 日宣布,其面向智能体 AI 工厂的下一代超级计算平台 NVIDIA Vera Rubin 已进入全面量产阶段。IT之家此前已有相关报道。 除此之外,英伟达同时确认新一代 Spectrum-X 以太网硅光技术已同步进入全面量产阶段,这是该平台实现大规模 AI 工厂网络互联的核心基石。 作为全球首款基于光电一…

AI 点评 · 硅光技术量产突破,能效提升5倍,将加速AI工厂网络部署,改变行业格局。

IT之家
行业信号
6/2 11:56
CPU 需求与日俱增,英特尔陈立武自曝许多公司 CEO 来电“求供货”

IT之家 6 月 2 日消息,据澎湃新闻,英特尔 CEO 陈立武 2 日(今天)在台北电脑展上表示,CPU 需求越来越高,但供给受到限制。过去四周内, 许多公司 CEO 打电话给他要更多的 CPU ,对英特尔来说“是一个机会”。 AI 智能体的兴起,使中央处理器的重要性得以再次提升,从而带动需求大量增加。陈立武在谈到 CPU 的发展趋势时指出,AI 智能体需…

AI 点评 · 高管亲述供货紧张,反映AI时代CPU需求爆发,英特尔产能成关键变量。

IT之家
行业信号
6/2 11:23
Rehumanizing global health care with agentic AI

The global health care sector is under increasing strain. Decades of chronic underinvestment and constraints in recruitment have coincided with a surge in demand for services for a…

AI 点评 · 用AI代理重构医疗流程,缓解人力短缺,提升服务效率与可及性。

MIT Tech Review
设计 / 产品
6/2 10:50
slavaZim/episodiq

Economical human-readable logs and structural (event pattern) retrieval for agentic trajectories

GitHub
设计 / 产品
6/2 09:50
腾讯客服:微信正与华为、荣耀、小米、OPPO、vivo 等合作,通过手机语音助理发起音视频通话或向指定好友发送消息

IT之家 6 月 2 日消息,据IT之家小伙伴今日反馈,腾讯客服最新回复显示, 微信正在与华为、荣耀、小米、OPPO、vivo 等手机厂商合作推出 A2A 助手能力 。 用户可以通过手机语音助理发起微信音视频通话或向指定好友发送消息。该功能基于 A2A(Agent-to-Agent)协作机制, 由厂商 AI 助手向微信发起指令,微信负责执行并返回结果 ,全程…

AI 点评 · 手机厂商AI助手与微信深度打通,标志着跨应用智能协作进入实用阶段。

IT之家
模型 / Agent
6/2 02:00
NVIDIA Jetson Brings Agentic AI to the Physical World

Agentic AI is getting physical. At COMPUTEX on Tuesday, NVIDIA announced NVIDIA JetPack 7.2 and NVIDIA NemoClaw support on NVIDIA Jetson. JetPack 7.2 brings agentic AI skills, Yoct…

AI 点评 · 英伟达让AI从虚拟走向实体,开启物理世界自主决策新纪元。

NVIDIA
设计 / 产品
6/1 22:38
阿里发布 Qwen3.7-Plus 模型,升级多模态交互混合 AI 智能体

IT之家 6 月 2 日消息,阿里千问大模型今天(6 月 2 日)发布博文,宣布推出 Qwen3.7-Plus 模型, 定位为多模态交互混合智能体。 Qwen3.7-Plus 是 Qwen3.7 的多模态升级版,核心定位是视觉与语言统一的智能体基座。 它保留文本、编码、工具使用和生产力工作流能力,同时强化视觉理解、视觉推理和跨模态任务处理。 模型已通过阿里云…

AI 点评 · 多模态与智能体融合,或加速AI从“对话”迈向“行动”的关键一步。

IT之家
Skill / 资源
6/1 21:31
OpenAI models and Codex on Amazon Bedrock are now generally available

GPT-5.5, GPT-5.4, and Codex are now generally available on Amazon Bedrock. Deploy them in production applications and agents today, on Bedrock’s high performance inference engine.

AI 点评 · OpenAI模型登陆亚马逊云平台,企业应用部署门槛进一步降低。

AWS ML
行业信号
6/1 20:00
Gemini’s new AI agent is about as good as Google’s demo

Google's new "24/7" AI agent, Gemini Spark, can be shockingly good at doing things on your behalf. But I'm not sure it's worth the financial cost and potential privacy tradeoffs. T…

AI 点评 · AI助手能力接近演示效果,但隐私与成本的双重代价仍需权衡。

The Verge
设计 / 产品
6/1 18:20
crimeacs/auto-improve

GAN-style self-improvement loop for any text artifact: mutate, grade with a SEPARATE model, keep only verified wins (pairwise-judged), revert the rest. The git…

GitHub
论文 / 方法
6/1 17:56
ClinEnv: An Interactive Multi-Stage Long Horizon EHR Environment for Agents

Clinical practice is not the selection of an answer from enumerated options: a physician gathers heterogeneous information incrementally and commits to sequential, irreversible decisions under uncerta…

AI 点评 · 电子健康记录多阶段交互环境,弥合了AI临床决策与真实医疗流程间的鸿沟。

arXiv
论文 / 方法
6/1 17:51
HERO'S JOURNEY: Testing Complex Rule Induction with Text Games

We introduce HERO'S JOURNEY, a benchmark for rule induction in goal-directed episodic tasks, where agents must infer hidden rules from demonstrations and act on them through multi-step execution. HERO…

AI 点评 · 用文本游戏测试AI规则归纳能力,填补了复杂推理任务基准的空白。

arXiv
论文 / 方法
6/1 17:45
SkillHarm: Lifecycle-Aware Skill-Based Attacks via Automated Construction

Agent skills occupy a privileged position in the agent workflow, as agents are expected to implicitly follow and execute them, rendering third-party skills a vulnerable attack surface. Existing studie…

AI 点评 · 自动化构建技能生命周期攻击,揭示第三方技能在智能体流程中的隐蔽安全风险,需重视防御。

arXiv
论文 / 方法
6/1 17:40
Tracking the Behavioral Trajectories of Adapting Agents

Text files such as skill files, memory files, and behavioral configuration files play a central role in defining how modern agents act. Through edits by humans or the agents themselves, these files ma…

AI 点评 · 追踪智能体行为轨迹,揭示自我调整机制,为AI决策透明化提供新视角。

arXiv
论文 / 方法
6/1 17:36
Auditing Asset-Specific Preferences in Financial Large Language Models: Evidence from Bitcoin Representations and Portfolio Allocation

Large language models now power robo-advisors and trading agents, yet whether they carry built-in biases toward specific assets is largely untested. We ask three questions: do LLMs systematically pref…

AI 点评 · 审计金融大模型对特定资产的偏好,揭示AI决策的隐性偏差,影响投资策略可靠性。

arXiv
Skill / 资源
6/1 16:12
AgentOps: Operationalize agentic AI at scale with Amazon Bedrock AgentCore

When you build agentic AI solutions, you face unique operational challenges. Agents make unpredictable decisions, costs spiral unexpectedly, and debugging non-deterministic failure…

AI 点评 · 亚马逊Bedrock AgentCore让AI代理规模化运营更可控,破解成本与调试难题。

AWS ML
行业信号
6/1 08:16
华为 FreeClip 2 耳夹耳机典藏版发布:珠宝盒设计、全新 AI 键智能体交互,1499 元

IT之家 6 月 1 日消息,在今天的华为 nova 16 系列及全场景新品发布会上,华为终端 BG CEO 何刚正式发布了 FreeClip 2 耳夹耳机典藏版, 定价 1499 元 。 据介绍,华为 FreeClip 2 耳夹耳机典藏版采用鎏光宝盒 + 珠宝盒设计,充电舱采用真空镀膜工艺,主打“圆润璀璨”, 同时内部空间提升 20% 。 这款耳机还与周大…

AI 点评 · 将珠宝美学与AI智能体交互结合,为耳机品类带来轻奢体验与技术创新突破。

IT之家
行业信号
6/1 05:00
“全球最强大的桌面 AI 超级计算机”,英伟达 DGX Station for Windows 发布

IT之家 6 月 1 日消息,在今日的 2026 台北国际电脑展主题演讲中,英伟达 CEO 黄仁勋发布了“全球最强大的桌面 AI 超级计算机”—— DGX Station for Windows 。 DGX Station for Windows 用于在 Windows 上开发和运行智能体 —— 基于英伟达 GB300 Grace Blackwell Ult…

AI 点评 · 首次将企业级AI算力带入桌面端,为Windows生态开发者提供了本地化训练与推理的超级工具。

IT之家
模型 / Agent
6/1 04:46
英伟达发布 5500 亿参数 Nemotron 3 Ultra 开源模型,较同级别前沿模型推理速度最高提升 5 倍

IT之家 6 月 1 日消息,为加强自主智能体的智能能力,英伟达今日发布了面向全天候运行智能体的全新开源模型与数据集,相关成果由英伟达 Nemotron 联盟联合打造。 据官方介绍,英伟达 Nemotron 3 Ultra 是一款拥有 5500 亿参数的混合专家模型,可为代码开发、科研及企业业务流程中的长效智能体提供顶尖智能能力。相较于同级别主流开源前沿模型…

AI 点评 · 参数规模与推理速度双突破,为智能体部署树立新标杆。

IT之家
模型 / Agent
6/1 04:30
NVIDIA Levels Up Local AI Agents Across RTX PCs and DGX Spark

Personal agents are exploding in popularity, with open source projects like OpenClaw and Hermes seeing rapid adoption by AI developer communities on GitHub. Built to adapt to indiv…

AI 点评 · 英伟达将本地AI智能体部署到RTX电脑和DGX工作站,推动个人AI应用从云端走向本地化。

NVIDIA
行业信号
6/1 04:23
英伟达 Vera 处理器发布:专为 AI 智能体打造,OpenAI、SpaceXAI、字节跳动都要用

IT之家 6 月 1 日消息,在今日的 2026 台北国际电脑展主题演讲中,英伟达 CEO 黄仁勋宣布正式推出 Vera 处理器 。 英伟达 Vera 是一款专为 AI 智能体打造的 CPU ,速度比 x86 处理器快 1.8 倍,可驱动各行各业的多样化工作负载,Vera 现已全面投产。 Vera 以 Grace CPU 的成功为基础(迄今为止,Grace…

AI 点评 · 巨头下场定义AI智能体专用芯片,生态号召力预示行业新标杆。

IT之家
设计 / 产品
6/1 03:55
黄仁勋:英伟达下一代 AI 超级芯片平台 Vera Rubin 全面投产

IT之家 6 月 1 日消息,在今日的 2026 台北国际电脑展主题演讲中,英伟达 CEO 黄仁勋宣布 Vera Rubin 全面投产。 Vera Rubin 为下一代 AI 工厂提供了 POD 规模的基础架构 —— 与上一代 Grace Blackwell 平台相比, 其大规模智能体吞吐量提高了 10 倍 。 凭借成熟的开源 MGX 设计,英伟达供应链生态…

AI 点评 · 下一代AI算力跃升10倍,英伟达再次定义超大规模集群新标杆。

IT之家
模型 / Agent
6/1 03:36
MiniMax M3 正式发布:前沿 Coding 能力、1M 上下文、原生多模态

MiniMax M3 今日正式发布。 MiniMax M3 在编程和智能体等专业任务上达到了前沿的能力。它使用了全新注意力架构 MSA (MiniMax Sparse Attention),最高支持 1M 超长上下文。它也是一个原生多模态模型,支持图片和视频的输入,并能操作电脑桌面。 在衡量 Coding 能力的 SWE-Bench Pro 上,MiniMa…

开源中国
设计 / 产品
6/1 03:20
couragec/llm-intern-skill

LLM internship resume and job-search Codex Skill: resume polish, JD tailoring, evidence guard, interview grilling, and Project Scout for LLM/RAG/Agent roles. 大模…

GitHub
设计 / 产品
6/1 03:20
couragec/LLMInternSkill

LLMInternSkill: LLM internship resume and job-search Codex Skill for resume polish, JD tailoring, evidence guard, interview grilling, and Project Scout. 大模型实习简历…

GitHub
设计 / 产品
6/1 02:29
RuleGo v0.36.0 发布:声明式 AI Agent 框架,规则引擎 × 智能体一体化

RuleGo 是一个基于 Go 语言的轻量级、高性能、嵌入式规则引擎。它通过规则链(JSON/可视化)编排组件,实现复杂业务逻辑的声明式管理,在物联网、边缘计算、数据集成、自动化等场景有广泛应用。 v0.36.0 是一个里程碑版本:rulego-components-ai 从 AI 组件库正式升级为声明式 AI Agent 开发框架,同时 Server 模块…

开源中国
模型 / Agent
5/31 23:46
GodeX 1.0.0 发布,面向 Codex 的 Responses API 兼容网关

让每个模型都成为 Codex 引擎。 OpenAI 兼容的 Responses API 网关,让 Codex、CLI 工具和开发者 Agent 接入任意模型。 English Documentation · 中文文档 GodeX 让使用 OpenAI Responses API 的客户端,可以通过一个本地网关调用 DeepSeek、Xiaomi、MiniMa…

开源中国
论文 / 方法
5/31 20:00
Joint Agent Memory and Exploration Learning via Novelty Signals

In open-ended environments, exploration is fundamental for autonomous agents, yet current language model agents struggle with this. Effective exploration requires memory, but retaining raw interaction…

AI 点评 · 结合新颖信号统一记忆与探索,让语言模型在开放环境中自主发现未知。

HuggingFace Papers
论文 / 方法
5/31 20:00
Multi-Agent Computer Use

Computer use agents (CUAs) today are primarily deployed as single serial agents. This setup is suboptimal for complex long-horizon tasks that benefit from task decomposition, parallel execution, and c…

AI 点评 · 多智能体协作提升复杂长任务效率,突破单代理局限,值得关注。

HuggingFace Papers
论文 / 方法
5/31 20:00
K-BrowseComp: A Web Browsing Agent Benchmark Grounded in Korean Contexts

Frontier model evaluations are shifting from foundational capabilities (e.g., instruction following and reasoning) toward compositional, agentic ones, but Korean agentic benchmarks remain scarce. We i…

AI 点评 · 首个聚焦韩语场景的网页浏览智能体评测基准,填补了非英语环境下的评估空白。

HuggingFace Papers
设计 / 产品
5/31 19:50
ClaudioDrews/memory-os

A 7-layer memory operating system for Hermes Agent — persistent memory with Qdrant, structured facts, fabric recall, auto-curated wiki, and surgical context inj…

GitHub
设计 / 产品
5/31 16:42
RUC-NLPIR/Arbor

A generalist autonomous research agent — runs experiments, researches, and iteratively optimizes, autonomously.

GitHub
设计 / 产品
5/31 15:45
duncatzat/vigils

A local control plane for AI agents — see what they do, approve what matters, keep secrets out. Rust + Tauri + Chrome MV3.

GitHub
设计 / 产品
5/31 07:22
argahv/sisyphus-academica

Sisyphus Academica — The Research Paper Writing Army. 20+ agent swarm: 6 novelty engines, 10 adversarial reviewers, Humanizer-integrated writing, citation verif…

AI 点评 · 用20多个AI代理模拟学术生产链,挑战论文写作与评审的自动化边界。

GitHub
设计 / 产品
5/30 20:14
SumanD18/sentinel

Open-source observability and trust layer for AI agents: trace every step, score every output, catch hallucinations and runaway loops in real time. Self-hostabl…

GitHub
论文 / 方法
5/30 20:00
3DCodeBench: Benchmarking Agentic Procedural 3D Modeling Via Code

Procedural 3D modeling through code is emerging as a versatile paradigm, offering deterministic, engine-ready, and precisely editable assets that neural 3D generators inherently lack. Authoring such p…

AI 点评 · 首个用代码评估AI三维建模能力的基准,填补了程序化生成与神经渲染之间的测评空白。

HuggingFace Papers
论文 / 方法
5/30 20:00
SkillAdaptor: Self-Adapting Skills for LLM Agents from Trajectories

Large language model (LLM) agents increasingly rely on reusable external skills to solve long-horizon interactive tasks. Existing training-free skill adaptation pipelines usually update skills from fu…

AI 点评 · 用轨迹数据让LLM代理自动进化技能,免训练自适应方案突破长程任务瓶颈。

HuggingFace Papers
论文 / 方法
5/30 20:00
Agent Skills Should Go Beyond Text: The Case for Visual Skills

Reusable skills are a key mechanism for extending agent capabilities, allowing agents to accumulate experience and solve increasingly complex tasks. Yet most existing skill-learning methods store reus…

AI 点评 · 视觉技能弥补语言局限,让智能体在复杂任务中更高效积累经验。

HuggingFace Papers
论文 / 方法
5/30 20:00
Trust Region On-Policy Distillation

On-Policy Distillation (OPD) is a fundamental technique for efficient post-training of large language models (LLMs), with broad applications in agent learning, multi-task enhancement, and model compre…

HuggingFace Papers
设计 / 产品
5/30 16:20
AtomFlow-AI/MoleCode

Molecode presents molecules as code and enables LLMs to operate and reason on chemistry directly.

GitHub
设计 / 产品
5/30 11:53
prashar32/riskkernel

Deterministic cost / loop / time budgets · full observability · crash-resumable runs · human-approval gates · a memory you own. Self-hosted. Your keys. No telem…

AI 点评 · 将确定性成本、循环时间预算与可恢复运行结合,为AI安全执行提供新范式。

GitHub
论文 / 方法
5/29 20:00
FineVerify: Scaling Test-Time Compute with Fine-Grained Self-Verification for Agentic Search

Agentic search requires language model agents to explore many sources and answer complex information-seeking questions. Scaling test-time compute is a promising way to improve these agents, but curren…

AI 点评 · 用细粒度自验证扩展测试时计算,首次系统解决智能体搜索中的错误累积问题,为复杂信息检索提供可扩展方案。

HuggingFace Papers
论文 / 方法
5/29 17:57
Stateful Online Monitoring Catches Distributed Agent Attacks

Language models can find thousands of severe software vulnerabilities, and agents are increasingly being misused for cyberattacks. To avoid detection, attackers frequently distribute their misuse, spl…

AI 点评 · 分布式智能体攻击难追踪,状态监测实现实时阻断,提升AI安全防御新高度。

arXiv
设计 / 产品
5/29 17:25
zhnt/loushang

AI-native coding orchestration platform: unified multi-model agent runtime with stateful sessions, tool governance, and traceable delivery.

GitHub
论文 / 方法
5/29 17:00
Preference-Aware Rubric Learning for Personalized Evaluation

As Large Language Models (LLMs) evolve from general-purpose assistants to user-centric agents, personalization has become central to aligning model behavior with individual preferences, making the eva…

AI 点评 · 个性化评估框架创新,让大模型更懂用户,提升人机交互体验。

arXiv
行业信号
5/29 10:00
Adobe’s conversational AI agent is a mediocre design intern

AI image tools rarely make me feel like I'm part of the creative process. They are, after all, mostly designed so that people with no design experience can type in a few words and…

AI 点评 · 直击AI工具痛点:设计过程缺乏参与感,暴露当前技术局限。

The Verge
设计 / 产品
5/29 08:19
huawei-csl/KVarN

KVarN is a native vLLM KV-cache quantization backend for your agents: 3-5x more context, throughput above FP16, and FP16-level accuracy. Calibration-free, one f…

GitHub
设计 / 产品
5/29 07:25
StarTrail-org/PixelRAG

The end of web parsing. The beginning of scalable pixel-native search.

AI 点评 · 将网页解析转向像素级原生搜索,为多模态检索开辟全新路径。

GitHub
行业信号
5/28 21:24
The internet is being rebuilt for machines

As AI agents move from experiments to production, AWS, Cloudflare, and others are redesigning cloud infrastructure for a future dominated by machine-generated internet traffic inst…

AI 点评 · 云巨头正为AI时代重构网络,机器流量将主导未来,基础设施变革迫在眉睫。

TechCrunch
设计 / 产品
5/28 20:58
JoniMartin27/lookspan

Local-first observability dashboard for AI agents. MCP-native. Look at every span your agents emit.

GitHub
Skill / 资源
5/28 20:32
Evaluating Deep Agents using LangSmith on AWS

This post combines learnings from LangChain’s work on evaluating deep agents and Anthropic’s guide to demystifying evals for AI agents into a practical guide. In this post, you wil…

AI 点评 · 结合LangChain与Anthropic的评估经验,为复杂AI代理提供实用评测指南,填补行业方法论

AWS ML
行业信号
5/28 20:06
Asana acquires no-code agent-builder StackAI

Asana will incorporate StackAI into its growing suite of AI workflow tools.

AI 点评 · Asana收购无代码智能体构建工具,加速AI工作流布局,降低企业自动化门槛。

TechCrunch
论文 / 方法
5/28 20:00
Task-Focused Memorization for Multimodal Agents

Long-term memory is essential for multimodal agents to build coherent experience, accumulate world knowledge, and achieve continual learning. However, constructing effective memory goes beyond memory…

AI 点评 · 聚焦多模态智能体的长期记忆构建,突破传统记忆局限,实现持续学习与知识积累。

HuggingFace Papers
论文 / 方法
5/28 20:00
MineExplorer: Evaluating Open-World Exploration of MLLM Agents in Minecraft

Multimodal large language models (MLLMs) have shown strong capabilities in perception, reasoning, and action generation. However, their ability to sustain exploration in dynamic open worlds remains un…

AI 点评 · 首个用《我的世界》评估多模态大模型开放世界探索能力的基准,填补了该领域测试空白。

HuggingFace Papers
模型 / Agent
5/28 17:51
Claude Opus 4.8 is now available on AWS

This post covers Opus 4.8's improvements and practical guidance for AI engineers integrating the model into agentic systems and production inference workloads on Amazon Bedrock.

AI 点评 · Claude新模型登陆AWS,为AI工程化部署提供关键升级,值得开发者关注。

AWS ML
模型 / Agent一手源
5/28 12:00
Vibe gets to work.

The unified agent for long-horizon productivity and coding, launching with Work and Code modes. Plus, a new Vibe VS Code extension.

可信度 88交叉信源 1
Mistral AI
模型 / Agent
5/28 12:00
How Endava builds an agentic organization with Codex

Learn how Endava uses Codex to build an agentic organization, accelerating software delivery and reducing requirements analysis from weeks to hours.

AI 点评 · 恩达瓦用Codex将需求分析从周缩短到小时,展示了AI代理加速软件交付的实战价值。

OpenAI
设计 / 产品
5/28 10:45
2aronS/Duel-Agents

CLI, SDK, and IDE plugins for Duel Agents

AI 点评 · 用命令行工具和插件简化AI智能体开发,提升调试效率。

GitHub
设计 / 产品
5/28 09:03
Health-Yang/MineEcho

Local-first Memory OS for personal AI assistants with L0-L3 memory, Wiki++ knowledge, skill routing, and TokenLess context compression.

AI 点评 · 个人AI助手本地记忆系统,实现知识路由与无令牌压缩,突破云端依赖瓶颈。

GitHub
设计 / 产品
5/28 01:46
modelstudioai/cli

Official Model Studio CLI(阿里云百炼 CLI)built for AI Agent frameworks, exposing models, search, multimodal, and workflow capabilities as structured tool calls.

GitHub
Skill / 资源
5/27 23:44
sqlite AGENTS.md

sqlite AGENTS.md SQLite gained an AGENTS.md file five days ago - but it's not intended for their own development, it's presumably aimed at people who are pointing agents at the SQL…

AI 点评 · SQLite新增AGENTS.md,专为AI代理设计,体现数据库与智能工具融合新趋势。

Simon Willison
设计 / 产品
5/27 22:47
helloianneo/ian-xiaohei-illustrations

中文小黑怪诞正文配图生成 Skill | 16:9 白底手绘 | 少量红橙蓝批注 | Codex Skill

AI 点评 · 结合手绘与AI生成,打造独特怪诞视觉风格,创意与工具融合的趣味尝试。

GitHub
Skill / 资源
5/27 20:06
Building AI agents for business support using Amazon Bedrock AgentCore

In this post, we share how the AWS Generative AI Innovation Center (GenAIIC) collaborated with Works Human Intelligence (WHI) to build two AI agents using Amazon Bedrock AgentCore.…

AI 点评 · 用亚马逊Bedrock AgentCore构建商业AI助手,为企业自动化客服与流程优化提供可落地的技

AWS ML
论文 / 方法
5/27 20:00
LongDS-Bench: On the Failure of Long-Horizon Agentic Data Analysis

Real-world data analysis is inherently iterative, yet existing benchmarks mostly evaluate isolated or short interactive tasks, leaving agents' ability to track evolving analytical context over long ho…

AI 点评 · 长期自主数据分析基准揭示AI在持续追踪复杂分析进程中的关键短板。

HuggingFace Papers
论文 / 方法
5/27 20:00
Harness Updating Is Not Harness Benefit: Disentangling Evolution Capabilities in Self-Evolving LLM Agents

LLM agents are increasingly deployed as systems built around editable external harnesses, including prompts, skills, memories and tools, that shape task execution without changing model parameters. Ha…

AI 点评 · 揭示大模型进化本质:外部系统更新不等于模型能力提升,为自我进化智能体研究厘清关键概念。

HuggingFace Papers
Skill / 资源
5/27 18:00
Powering agentic AI sales strategy with Amazon Bedrock AgentCore

As agent adoption scaled, we saw a common pattern emerge across enterprises, including our own sales organization: specialized agents deliver value, but without orchestration, user…

AI 点评 · 用Bedrock AgentCore编排多智能体协作,是企业规模化部署AI销售的关键突破。

AWS ML
模型 / Agent
5/27 16:00
AI Factories: The New Infrastructure of Intelligence

AI factories are token factories, converting power into intelligence in real time. And as agentic AI scales and autonomous, always-on special agents are deployed in the enterprise,…

AI 点评 · AI工厂将电力实时转化为智能,标志着智能基础设施革命的开端。

NVIDIA
设计 / 产品
5/27 14:44
DYAI2025/Plumbline

Plumbline — a self-learning, customer-value-governed agile AI agent team for Claude Code. 87 subagents + skills, TDD defense-in-depth gates, Kaizen retros, a fo…

GitHub
设计 / 产品
5/27 12:46
kesslernity/awesome-copilot-chat-agents

82 ready-to-deploy Microsoft Copilot Chat agents — paste the instruction block into Copilot Studio and you're live. Writing, HR, PM, IT ops, Sales, Finance, Eng…

GitHub
设计 / 产品
5/27 12:05
op7418/guizang-social-card-skill

🪧 Claude Code / Codex skill — generate Xiaohongshu carousels & WeChat 21:9+1:1 cover pairs. Editorial × Swiss visual systems, 28 layouts, 10 themes, single-fil…

AI 点评 · 将小红书和微信封面设计自动化,融合瑞士视觉系统,极大提升内容生产效率。

GitHub
设计 / 产品
5/27 08:25
yb2460/harness-anything

Harness Anything - AI agent control hub: WPS, MS Office, Zotero, Photoshop, 47 CLI commands, 27 academic skills, SVG-to-PPTX

AI 点评 · 用命令行让AI自动操控WPS三大组件,打通办公软件自动化新路径。

GitHub
模型 / Agent
5/27 07:00
Building self-improving tax agents with Codex

See how OpenAI, Thrive, and Crete built a self-improving tax agent with Codex, automating filings, improving accuracy, and accelerating workflows.

AI 点评 · 利用Codex实现税务代理自我进化,自动化与准确性双提升,开辟AI落地新场景。

OpenAI
设计 / 产品
5/27 05:46
withkynam/vibecode-pro-max-kit

Your AI forgets. This remembers. Spec-driven coding harness for vibecoders, product owners, CEOs and real builders — self-improving context memory, 15 agents, 3…

AI 点评 · 用12个智能体构建自进化记忆系统,专为追求高效编码的实干者设计,重新定义AI协作体验。

GitHub
设计 / 产品
5/27 02:08
nexu-io/html-video

Programmatic video for coding agents — HTML to video on your laptop. Turn HTML, CSS & data into real MP4s with pluggable render engines, 21 templates, AI soundt…

GitHub
模型 / Agent
5/27 00:00
Warp’s big bet on building open source with GPT-5.5

Warp uses GPT-5.5 and OpenAI models to coordinate coding agents across local, cloud, and open-source development workflows.

AI 点评 · Warp结合GPT-5.5与开源,探索跨环境编程新范式,值得关注。

OpenAI
论文 / 方法
5/26 20:00
A Matter of TASTE: Improving Coverage and Difficulty of Agent Benchmarks

As agent capabilities advance, existing benchmarks, such as τ^2-Bench, are becoming increasingly saturated. Yet constructing new benchmark tasks remains complex, costly, and labor-intensive. Moreover,…

AI 点评 · 通过自动化生成更难更全的基准任务,突破现有评测瓶颈,为智能体能力评估提供新思路。

HuggingFace Papers
设计 / 产品
5/26 19:16
shyftlabs/continuum

Continuum — the agent runtime by ShyftLabs. Build, orchestrate, ship.

AI 点评 · ShyftLabs推出智能体运行时,简化构建到部署全流程,值得开发者关注。

GitHub
Skill / 资源
5/26 15:36
Microsoft Copilot Cowork Exfiltrates Files

Microsoft Copilot Cowork Exfiltrates Files The biggest challenge in designing agentic systems continues to be preventing them from enabling attackers to exfiltrate data. In this ca…

AI 点评 · 揭示AI安全短板:Copilot被利用外泄文件,警示企业需警惕智能助手的数据防护漏洞。

Simon Willison
行业信号
5/26 14:54
Rethinking organizational design in the age of agentic AI

Amid rapidly growing adoption of enterprise-level AI agents, there’s a disconnect emerging between ambition and execution. Although 85% of organizations say they want to be agentic…

AI 点评 · 企业级AI代理快速增长,组织架构面临颠覆性变革,平衡雄心与执行是关键看点。

MIT Tech Review
设计 / 产品
5/26 12:45
fancyboi999/ai-engineering-from-scratch-zh

Agent工程师最全学习路径 · 从零精通 AI 工程 · 20 阶段 503 课 · 中文全量翻译 + 配套站点 + 动画讲解视频 · 如何成为 AI Agent 工程师的修成指南

GitHub
设计 / 产品
5/26 07:45
biao994/DocPaws

工程化 RAG 文档助手:知识库、PDF 索引、Agent 工具编排、scope 检索、引用溯源与拒答阈值。FastAPI + Vue3

AI 点评 · 企业级RAG落地范本,从检索拒答到工具编排的完整工程化实践。

GitHub
设计 / 产品
5/25 20:26
Tejas-TA/predikit

The missing bridge between your ML models and your AI agents.

GitHub
设计 / 产品
5/25 11:06
UditAkhourii/adhd

ADHD — a skill for coding agents. Tree-of-thought with pruning, built on the Claude & Codex Agent SDK. Fans out parallel divergent thoughts under different cogn…

AI 点评 · 用树状思维加剪枝策略,让编码代理模拟多动症思考,提升复杂问题解决效率。

GitHub
设计 / 产品
5/25 00:05
oleksiijko/pmb

Local-first persistent memory for AI coding agents (Claude Code, Cursor, Codex) via MCP. 94.5% LoCoMo recall@10, 70ms p50, multilingual, zero API keys.

AI 点评 · 本地优先持久记忆方案,大幅提升AI编码代理效率,无需API密钥,性能指标出色。

GitHub
设计 / 产品
5/23 16:16
ongridio/ongrid

An ops AI Agent that understands your infrastructure, finds the root cause, and fixes it — right from Slack, Telegram, Lark or DingTalk.

GitHub
设计 / 产品
5/23 02:44
study8677/awesome-architecture

🧭 Architecture-first system design: 26 bilingual tutorials, 25 architecture templates, and 6 end-to-end cases covering distributed systems, AI-native systems,…

GitHub
设计 / 产品
5/22 08:31
leestott/foundry-cicd

Enterprise-ready CI/CD reference for Microsoft Foundry AI agents, with parallel GitHub Actions and Azure DevOps pipelines, evaluation-driven quality gates, and…

AI 点评 · 企业级AI代理的CI/CD参考方案,实现并行流水线与质量门控,提升部署效率与可靠性。

GitHub
模型 / Agent
5/22 00:00
OpenAI named a Leader in enterprise coding agents by Gartner

OpenAI is named a leader in the 2026 Gartner Magic Quadrant for Enterprise AI Coding Agents, with Codex recognized for innovation and enterprise-scale deployment.

AI 点评 · Gartner权威认证,OpenAI编码智能体在创新与规模化部署上领先行业。

OpenAI
设计 / 产品
5/21 13:58
mims-harvard/AutoScientists

AutoScientists: Self-Organizing Agent Teams for Long-Running Scientific Experimentation

AI 点评 · 自组织AI团队实现长期科学实验,推动自动化研究范式突破。

GitHub
设计 / 产品
5/21 11:14
wangchuxiaoji-oss/doubao2api

Reverse-engineered Doubao (豆包) API → OpenAI-compatible REST service. Free multimodal chat, image/video/music generation, and file hosting for AI agents.

AI 点评 · 逆向工程将豆包API转为OpenAI兼容接口,免费提供多模态功能,大幅降低AI开发门槛。

GitHub
设计 / 产品
5/21 03:58
Eynzof/Hermes-CN-Desktop

Hermes Agent CN desktop app, Windows-First, built with Tauri, Typescript and Rust. Isolated Hermes Agent core insides.

GitHub
设计 / 产品
5/20 19:14
NanoFlow-io/engram

🧠 Hybrid long-term memory plugin for OpenClaw agents — SQLite+FTS5 for structured facts, LanceDB for semantic recall

AI 点评 · 结合SQLite与向量数据库,为AI代理提供结构化事实与语义回忆的双重记忆支持。

GitHub
设计 / 产品
5/20 18:52
zhongweiv/hermes-edu-skills

中文教育 Agent Skill Pack:教材同步、备考复习、拍照答疑、错题复盘、亲子陪学、阅读写作和教师工具,Hermes Agent 可直接使用,也可导出到 OpenClaw/Codex/Cursor/Claude Code。

GitHub
设计 / 产品
5/20 03:24
VibeBench/VibeSearchBench

🔍 The hardest search benchmark in the wild — vague, multi-turn, proactive. 200 long-horizon tasks with persona-driven progressive disclosure, scored by verifia…

AI 点评 · 首个模糊多轮搜索基准,考验AI主动追问能力,填补了复杂意图检索评估的空白。

GitHub
模型 / Agent
5/19 17:45
I/O 2026: Welcome to the agentic Gemini era

The latest from Google I/O: See how we’re helping you get more done with Gemini.

AI 点评 · 谷歌发布Agentic Gemini,标志AI从工具向自主行动者进化,定义人机协作新范式。

Google AI
设计 / 产品
5/19 14:01
elvisun/newsjack

The open-source skills that turn your agent into a full PR team.

GitHub
设计 / 产品
5/19 11:59
ather-techie/rag-interview-questions

A comprehensive interview preparation guide covering all major RAG (Retrieval-Augmented Generation) architectures. 140 questions across 12 types, from Naive RAG…

GitHub
设计 / 产品
5/19 11:59
ather-techie/rag-interview-system

A complete collection of RAG interview questions, answers (286 questions & 18 RAG types), system design scenarios, architecture patterns, and production-ready c…

GitHub
设计 / 产品
5/19 09:44
langfuse/langfuse-workshop

End-to-end Langfuse workshop using a TypeScript Agent to teach the AI engineering loop: tracing, prompt management, monitoring, datasets, experiments, and evalu…

GitHub
设计 / 产品
5/18 23:04
JSingletonAI/dejavu

Memory that follows you across every AI tool. No cloud storage. No account required. Set it up once, use it everywhere.

AI 点评 · 打破工具壁垒的本地记忆系统,让AI实现跨平台无缝复用。

GitHub
模型 / Agent
5/18 21:48
Vera Arrives: NVIDIA’s First CPU Built for Agents Lands at Top AI Labs

The first NVIDIA Vera CPUs arrived at three of the world's leading AI labs on Friday — Anthropic in San Francisco, OpenAI in Mission Bay, SpaceXAI in Palo Alto — followed by a deli…

AI 点评 · 英伟达首款CPU专为AI代理设计,直供顶级实验室,或改写智能算力格局。

NVIDIA
行业信号
5/18 15:40
Show HN: InsForge – Open-source Heroku for coding agents

Hi HN, I'm Hang, cofounder of InsForge (YC P26). InsForge is an open-source Heroku for AI coding agents: a backend platform designed for coding agents to deploy, operate, and debug…

AI 点评 · 开源首个面向AI编码代理的Heroku式平台,填补了代理部署与调试的空白,值得开发者关注。

Hacker News
设计 / 产品
5/17 14:18
Second-Inc/second

The factory for custom internal software, purpose-built for human2agent work.

GitHub
设计 / 产品
5/17 04:33
openthomas-com/openthomas

Cut the cost of your agent fleet without switching agents. Makes Claude Code Dynamic Workflows cheap: the planner stays on Opus, the hundreds of parallel subage…

GitHub
设计 / 产品
5/16 17:44
sam-siavoshian/agent-notch

macOS computer-use agent in the notch. Long-press, talk, Claude drives the mouse.

AI 点评 · 把AI代理嵌入Mac刘海区域,长按语音操控鼠标,交互方式创新且实用。

GitHub
设计 / 产品
5/15 21:32
DenisSergeevitch/agents-best-practices

Provider-neutral Agent Skill for Codex, Claude Code, and agentic harness design.

AI 点评 · 通用Agent技能框架,适用于多种主流AI编码工具,提升开发效率与互操作性。

GitHub
设计 / 产品
5/15 09:19
husu/loom

一个写接口文档的AI Agent。支持使用Vibe coding 的方式,编写接口文档,同时自带友好的文档查看工具与接口Mock工具

GitHub
设计 / 产品
5/15 08:50
fangwendongcs/Auto-agent-factory

A production-ready toolkit to accelerate and automate the end-to-end lifecycle of AI Agent development.

AI 点评 · 助力企业快速部署AI代理,填补了开发到生产的工具链空白。

GitHub
设计 / 产品
5/15 06:59
CONSTELLATION-ENGINE/constellation-engine

Most AI agents forget you the moment the tab closes. Constellation Engine gives them a hippocampus — a living star map with spreading activation, Hebbian writeb…

GitHub
设计 / 产品
5/15 05:00
intellicia-public/parastore

Draw a store, generate LLM personas, and watch them shop — an isometric 3D sandbox for synthetic-consumer experiments.

GitHub
设计 / 产品
5/14 11:09
Purewhiter/mobilegym

MobileGym: A Verifiable and Highly Parallel Simulation Platform for Mobile GUI Agent Research · 浏览器里运行的安卓模拟器 · Browser-hosted Android Simulator · Verifiable Eva…

AI 点评 · 移动端GUI智能体研究提速,浏览器运行安卓模拟器实现可验证并行测试。

GitHub
设计 / 产品
5/12 19:31
gi-dellav/zerostack

Minimal coding agent written in Rust, optimized for memory footprint and performance

AI 点评 · 用Rust打造极简编码代理,专注内存优化与性能,为轻量化AI工具开辟新路径。

GitHub
设计 / 产品
5/12 13:07
johunsang/semble_rs

Fast, AI-agent-native code search in Rust — hybrid BM25 + semantic, Tree-sitter AST chunking, dependency & impact analysis. Drop-in replacement for grep/cat/rea…

AI 点评 · 用Rust实现的高性能AI原生代码搜索,结合混合检索与依赖分析,有望替代传统工具。

GitHub
设计 / 产品
5/12 04:14
AzmxAI/azmx

AZMX AI — The sovereign agent platform.

GitHub
设计 / 产品
5/11 22:55
secureagentics/Adrian

Runtime security monitoring and control for AI agents. Catches malicious tool use, prompt injection, and policy drift in real time, before the agent acts.

GitHub
设计 / 产品
5/11 21:18
sparkplug604/praxis

Local-first RAG and agent skills framework for source-traceable agent memory.

AI 点评 · 本地优先架构让RAG技能框架实现源头可追溯,为AI代理记忆管理提供新范式。

GitHub
设计 / 产品
5/11 11:19
juanjuandog/FinSight-AI

AI equity research agent with resilient workflows, Redis Lua single-flight, pgvector RAG, versioned reports, evidence tracing, and RAG evaluation.

AI 点评 · 高效AI投研工具,结合弹性工作流与证据溯源,提升研报可信度。

GitHub
设计 / 产品
5/11 09:40
nexu-io/html-anything

✨ The agentic HTML editor — your local AI agent writes the HTML, you ship it. 🚀 75 Skills × 9 Surfaces (magazine · deck · poster · XHS / tweet · prototype · da…

AI 点评 · 本地AI代理直接生成可交付的HTML,覆盖多种设计场景,大幅降低前端开发门槛。

GitHub
设计 / 产品
5/10 19:42
Kaelio/ktx-ai-data-agents-context

ktx is an executable context layer for data and analytics agents 🐙 Allow Claude Code, Codex, and any AI agent to query data accurately through MCP with skills,…

GitHub
设计 / 产品
5/10 19:42
Kaelio/ktx

ktx is an executable context layer for data and analytics agents 🐙 Allow Claude Code, Codex, or any other AI agent to query data accurately and with full conte…

AI 点评 · 用MCP技能层让AI代理精准查询数据,打通代码与分析的执行壁垒。

GitHub
设计 / 产品
5/10 12:29
namphuongtran/awesome-ai-coding-agent-tools

A curated list of tools, libraries, MCP servers, and frameworks that power AI coding agents.

AI 点评 · 资源聚合清单,帮你快速找到提升AI编程效率的利器。

GitHub
设计 / 产品
5/10 01:47
LichAmnesia/openseek

OpenSeek - 广度求索: open-source TUI coding agent with multi-provider routing, MCP, LSP, and Plan/Agent/YOLO modes.

GitHub
设计 / 产品
5/9 18:10
beltromatti/get-it

Read it. See it. Get it. Built at GDG AI Hack Milan 2026 for "Learn Different" track.

GitHub
设计 / 产品
5/9 17:49
byte5ai/omadia

Self-hostable agentic OS — build, run & audit multi-agent AI teams from signed plugins. Bring your own LLM key, own all your data, EU/GDPR-ready.

GitHub
设计 / 产品
5/9 10:22
recomby-ai/recomby-geo

GEO 领域 AI 员工开源方案 · Open-source GEO AI-employee solution (MIT). GEO Skills package + curated lists of agents and office CLIs that make up the AI-employee stack.

GitHub
设计 / 产品
5/9 06:16
tophant-ai/promptbeat

Break your AI before they do.

AI 点评 · 红队测试工具,主动发现AI模型安全漏洞,强化防御。

GitHub
设计 / 产品
5/9 04:46
ngaut/agent-git-service

Reimplement GitHub for Agents.

AI 点评 · 用Rust重写GitHub服务,专为AI代理设计,或开启自动化协作新范式。

GitHub
设计 / 产品
5/9 04:39
sno-ai/llmix

Production LLM call layer for AI agents and tools: keep OpenAI/Anthropic/AI SDK/LiteLLM, hot-swap models with MDA presets, and add cache, retries, circuit break…

AI 点评 · 统一多模型调用层,提升AI代理的稳定性和灵活性,降低开发成本。

GitHub
设计 / 产品
5/8 18:41
finewood2008/centaur-loop

半人马环 Centaur Loop:面向 AI Agent 反馈闭环、人类治理和记忆复盘的开源工作台 / Human-governed AI feedback loop workbench.

AI 点评 · 开源AI Agent工作台,打通人类治理与反馈闭环,助力记忆复盘,实用价值高。

GitHub
设计 / 产品
5/8 18:41
finewood2008/centaurloop

半人马环 Centaur Loop:AI 员工的最小工作单元框架。把复杂岗位拆解为可由 AI 接管、由人类治理、由反馈和记忆持续进化的循环工作流 / The smallest work unit for building AI employees.

AI 点评 · 开源AI治理工具,填补了Agent闭环管理的空白,兼顾人类监督与记忆复盘,实用性很强。

GitHub
设计 / 产品
5/8 10:03
stormzhang/token-tracker

Track token usage across local AI agents (Claude Code, Codex) — Custom StatusLine, CLI Dashboard with cost analysis, rate limit monitoring, and session tracking

GitHub
设计 / 产品
5/8 07:37
haydenbleasel/files-sdk

A unified storage SDK for object and blob backends. One small, honest API. Web-standards I/O.

AI 点评 · 统一存储SDK实现对象与二进制后端兼容,简化开发流程,值得关注。

GitHub
设计 / 产品
5/8 06:57
volcengine/SearchCLI

Open CLI for integrating AI search, recommendation, and conversational retrieval into agent systems and business systems

AI 点评 · 用命令行整合AI搜索与推荐,降低智能系统集成门槛,提升开发效率。

GitHub
设计 / 产品
5/7 23:22
zendev-sh/zenflow

Multi-agent orchestration & workflow engine. Declarative YAML workflows, LLM coordinator with hub-and-spoke mailboxes, race-safe delivery. One YAML file, one Go…

AI 点评 · 用声明式YAML编排多智能体工作流,结合LLM协调与安全投递,降低开发门槛。

GitHub
设计 / 产品
5/7 14:18
freestylefly/wesight

Open-source desktop AI agent workspace with one-click Claude Code, Codex, OpenClaw, Hermes Agent setup and custom LLM model routing.

AI 点评 · 开源桌面AI工作区整合多模型一键部署,降低智能体开发门槛,推动个性化工具构建。

GitHub
设计 / 产品
5/6 17:43
opensquilla/opensquilla

OpenSquilla — Token-Efficient AI Agent with same budget, higher intelligence density

AI 点评 · 用更少token实现更高智能密度,开源AI Agent效率突破值得关注。

GitHub
设计 / 产品
5/6 11:12
OpenOSINT/OpenOSINT

AI-powered OSINT agent with interactive REPL, MCP server, and CLI. 16 tools. Works with Claude, GPT-4, or local models. For authorized security research only.

AI 点评 · 开源AI驱动的OSINT工具,整合交互式命令行与多模型支持,为安全研究提供高效情报分析能力。

GitHub
设计 / 产品
5/5 21:08
NirDiamant/Agent_Memory_Techniques

Agent memory for LLMs: 30 runnable Jupyter notebooks covering conversation buffers, vector stores, knowledge graphs, episodic and semantic memory, MemGPT, Mem0,…

AI 点评 · 30个可运行笔记系统梳理LLM记忆机制,实操价值高,覆盖从基础到前沿。

GitHub
设计 / 产品
5/5 17:55
agynio/platform

Agyn is an open-source Kubernetes-native runtime that moves AI agents like Claude Code and Codex from laptops to company infrastructure with the controls enterp…

AI 点评 · 开源Kubernetes原生方案,让企业安全托管AI代理,填补了从个人工具到平台级部署的空白。

GitHub
设计 / 产品
5/5 16:30
yuc16/PatentRadar

自动化专利侵权竞品分析系统 —— 输入专利公开号,1 小时产出律师可复核的 claim chart 报告(逐特征对比 + 证据URL + 下一步建议);同时打包成 skill,可被任意 agent 调用。

AI 点评 · 专利侵权分析自动化,律师级报告1小时生成,大幅提升IP尽调效率。

GitHub
设计 / 产品
5/4 13:43
ybuild-ai/ai-game-art-pipeline-skill

Agent skill for turning AI images and videos into playable game art assets

AI 点评 · 聚焦AI图像转游戏资产的自动化流程,大幅降低游戏开发门槛。

GitHub
设计 / 产品
5/4 09:14
jmerelnyc/Photo-agents

Autonomous self-evolving agents. Vision-grounded layered memory and self-written skills for LLM agents that operate your computer.

AI 点评 · 自主进化代理结合视觉记忆,让AI真正学会操作电脑,突破传统指令限制。

GitHub
设计 / 产品
5/3 11:04
shawn0728/OpenSearch-VL

🔍 OpenSearch-VL provides a fully open recipe for training strong multimodal deep search agents through high-quality data curation, diverse visual/search tools,…

GitHub
设计 / 产品
5/3 08:16
pingchesu/hermes-curator-evolver

Evidence-driven skill evolution for Hermes Agent — reports, dry-run proposals, candidate search, and guarded apply

GitHub
Skill / 资源
4/4 11:45
Components of A Coding Agent

How coding agents use tools, memory, and repo context to make LLMs work better in practice

AI 点评 · 拆解编码代理三大核心模块,为LLM落地提供实用框架。

可信度 74交叉信源 1
Sebastian Raschka
模型 / Agent一手源
3/23 16:00
Speaking of Voxtral

Voxtral TTS: A frontier, open-weights text-to-speech model that’s fast, instantly adaptable, and produces lifelike speech for voice agents.

可信度 88交叉信源 1
Mistral AI
模型 / Agent一手源
7/22 13:00
Qwen3-Coder: Agentic Coding in the World

GITHUB HUGGING FACE MODELSCOPE DISCORD Today, we’re announcing Qwen3-Coder, our most agentic code model to date. Qwen3-Coder is available in multiple sizes, but we’re excited to in…

可信度 88交叉信源 1
通义千问 Qwen
Skill / 资源
11/28 00:00
Reward Hacking in Reinforcement Learning

Reward hacking occurs when a reinforcement learning (RL) agent exploits flaws or ambiguities in the reward function to achieve high rewards, without genuinely learning or completin…

AI 点评 · 强化学习易钻空子,揭示AI安全核心挑战,关乎真实任务可靠性。

Lilian Weng
Skill / 资源
11/28 00:00
Reward Hacking in Reinforcement Learning

Reward hacking occurs when a reinforcement learning (RL) agent exploits flaws or ambiguities in the reward function to achieve high rewards, without genuinely learning or completin…

可信度 74交叉信源 1
Lilian Weng
Skill / 资源
6/23 00:00
LLM Powered Autonomous Agents

Building agents with LLM (large language model) as its core controller is a cool concept. Several proof-of-concepts demos, such as AutoGPT , GPT-Engineer and BabyAGI , serve as ins…

AI 点评 · 揭示大语言模型作为核心控制器,推动自主智能体从概念走向实用化,标志AI应用新里程碑。

Lilian Weng
Skill / 资源
6/23 00:00
LLM Powered Autonomous Agents

Building agents with LLM (large language model) as its core controller is a cool concept. Several proof-of-concepts demos, such as AutoGPT , GPT-Engineer and BabyAGI , serve as ins…

可信度 74交叉信源 1
Lilian Weng