每日简报

2026-09-18

← 历史归档

alibaba/open-code-review

Go · ★ 34,625 · 🍴 2,461 · 📈 3,290 stars today

Fast, efficient, battle-tested at Alibaba's scale. Hybrid architecture code review tool: deterministic pipelines + LLM Agent, precise line-level comments, built-in multi-language ruleset (NPE, thread-safety, XSS, SQL injection), OpenAI & Anthropic compatible.

中文介绍 阿里开源的代码审查工具,采用混合架构,结合确定性管道和LLM代理,提供精确的行级注释和内置的多语言规则集,适用于大规模代码审查。

cloudflare/security-audit-skill

JavaScript · ★ 10,496 · 🍴 561 · 📈 3,606 stars today

A coding-agent skill for multi-phase security audits with independently verified, machine-readable findings

中文介绍 Cloudflare开发的代码安全审计技能,用于多阶段安全审计,提供经过独立验证的机器可读结果。

addyosmani/agent-skills

JavaScript · ★ 95,813 · 🍴 10,144 · 📈 680 stars today

Production-grade engineering skills for AI coding agents.

中文介绍 为AI编码代理提供生产级工程技能的库。

Tencent/BrowserSkill

TypeScript · ★ 4,070 · 🍴 288 · 📈 1,350 stars today

Let AI agents use your real, logged-in browser without interrupting your work. CLI + extension for browser automation across any shell-capable AI agent.

中文介绍 腾讯开发的浏览器自动化工具,允许AI代理使用真实登录的浏览器,而不会打断用户的工作。

alphaXiv/OpenResearch

Rust · ★ 4,915 · 🍴 304 · 📈 940 stars today

Turn your coding agents into research agents

中文介绍 将编码代理转变为研究代理的工具。

anthropics/claude-code

TypeScript · ★ 145,845 · 🍴 23,616 · 📈 538 stars today

Claude Code is an agentic coding tool that lives in your terminal, understands your codebase, and helps you code faster by executing routine tasks, explaining complex code, and handling git workflows - all through natural language commands.

中文介绍 Claude Code是一个终端中的代理编码工具,理解代码库,通过执行常规任务、解释复杂代码和处理git工作流来帮助用户更快地编码。

NationalSecurityAgency/ghidra

Java · ★ 78,445 · 🍴 8,679 · 📈 912 stars today

Ghidra is a software reverse engineering (SRE) framework

中文介绍 Ghidra是一个软件逆向工程框架,用于分析、修改和重用软件。

anthropics/knowledge-work-plugins

Python · ★ 24,541 · 🍴 2,939 · 📈 287 stars today

Open source repository of plugins primarily intended for knowledge workers to use in Claude Cowork

中文介绍 Claude Cowork知识工作插件开源库,主要为知识工作者提供使用。

Tencent/WeKnora

Go · ★ 26,155 · 🍴 3,548 · 📈 1,123 stars today

Open-source LLM knowledge platform: turn raw documents into a queryable RAG, an autonomous reasoning agent, and a self-maintaining Wiki.

中文介绍 腾讯开源的LLM知识平台,将原始文档转换为可查询的RAG、自主推理代理和自维护Wiki。

abue-ammar/tinycast

Swift · ★ 6,124 · 🍴 286 · 📈 738 stars today

Tinycast — a tiny, fully native macOS launcher, hotkeys, and clipboard history.

中文介绍 Tinycast是一个小巧的macOS启动器,支持热键和剪贴板历史记录。

cilium/cilium

Go · ★ 25,251 · 🍴 4,073 · 📈 153 stars today

eBPF-based Networking, Security, and Observability

中文介绍 基于eBPF的Networking、Security和Observability解决方案。

jamiepine/voicebox

TypeScript · ★ 54,818 · 🍴 6,823 · 📈 665 stars today

The open-source AI voice studio. Clone, dictate, create.

中文介绍 开源的AI语音工作室,支持克隆、语音输入和创建内容。

affaan-m/ECC

JavaScript · ★ 261,116 · 🍴 39,085 · 📈 1,173 stars today

The agent harness performance optimization system. Skills, instincts, memory, security, and research-first development for Claude Code, Codex, Opencode, Cursor and beyond.

中文介绍 Claude Code等代理的性能优化系统,提供技能、本能、记忆、安全和研究优先的开发。

roboflow/supervision

Python · ★ 50,789 · 🍴 4,828 · 📈 327 stars today

We write your reusable computer vision tools. 💜

中文介绍 Roboflow提供可重用的计算机视觉工具。

JustVugg/colibri

C · ★ 35,720 · 🍴 3,748 · 📈 872 stars today

Run frontier MoE models on hardware you already own — pure C, zero deps, experts streamed from disk. Tiny engine, immense model. 🐦

中文介绍 运行前沿MoE模型的工具,纯C语言编写,无需依赖,模型专家从磁盘流式传输。

TencentCloud/Octop

Python · ★ 3,408 · 🍴 352 · 📈 386 stars today

A smarter, self-hosted AI assistant — multi-user, multi-agent.

中文介绍 腾讯云的智能AI助手,支持多用户和多代理。

ever-co/ever-gauzy

TypeScript · ★ 7,511 · 🍴 1,119 · 📈 469 stars today

Ever® Gauzy™ - Open Business Management Platform (ERP/CRM/HRM/ATS/PM) - https://gauzy.co

中文介绍 Ever Gauzy是一个开放的业务管理平台,提供ERP、CRM、HRM、ATS和PM等功能。

cline/cline

TypeScript · ★ 68,541 · 🍴 7,413 · 📈 381 stars today

Autonomous coding agent as an SDK, IDE extension, or CLI assistant.

中文介绍 作为SDK、IDE扩展或CLI助手的自主编码代理。

coder/coder

Go · ★ 14,805 · 🍴 1,485 · 📈 204 stars today

Secure environments for developers and their agents

中文介绍 为开发者和他们的代理提供安全的环境。

n8n-io/n8n

TypeScript · ★ 204,943 · 🍴 60,742 · 📈 319 stars today

Fair-code workflow automation platform with native AI capabilities. Combine visual building with custom code, self-host or cloud, 400+ integrations.

中文介绍 具有原生AI能力的公平代码工作流自动化平台,结合可视化构建和自定义代码,支持自托管或云部署,拥有400多个集成。

CERA-MoA: Co-Evolving Routing Mechanisms with Continually Learning LLM Agents

👍 4

Current Mixture-of-Agents (MoA) paradigms generally treat query routing and agent fine-tuning as separate processes, limiting their ability to respond to evolving agent capabilities. This disconnect prevents routing strategies from adapting to evolving agent capabilities during post-training and pre

In-Context Robot Learning with VLM Agents

👍 14

Enabling robots to adapt to unfamiliar environments as readily as humans remains a moonshot goal of embodied AI. No finite collection of demonstrations can cover every task and situation a robot will encounter, making the ability to learn from context at deployment essential for generalization. Such

PANORAMA: Panoptic Grounded Captioning via Mask Proposal Selection

👍 9

Intelligent systems that act in the world require image understanding that is both comprehensive and spatially grounded. Current vision-language models (VLMs) can generate fluent and detailed image captions, but reliably associating them with image pixels remains challenging. Existing methods that c

A Zeroth-Order Paradigm for LLM Preference Alignment

👍 22

Direct preference alignment methods are widely used to align large language models (LLMs) with human preferences because of their computational and memory efficiency. However, likelihood displacement motivates alternative ways to extract information from preference pairs with small likelihood margin

Agora: Git as Shared Memory for Collective AutoResearch

👍 39

Autonomous research loops such as AutoResearch show that one coding agent can improve a training setup unattended. Run several of them and each session starts from scratch, so more agents tend to mean more duplicated search rather than more discovery. Agora is a shared memory for such agents: resear

Fathom: Per-Query Read Depth for Sparse Decoding over Offloaded KV Caches

👍 2

When agentic sessions run to a million tokens with many sessions resident at once, the KV cache and the index that ranks it live in host memory, and the scan that ranks all n keys for a top-k step becomes the traffic that bounds decoding. We present Fathom, a key scan in which each query decides how

ImpossibleRubrics: Stress-Testing Generated Rubrics as Reward Signals

👍 3

Language model-generated rubrics are increasingly used as reward signals for rubric-based reinforcement learning, LLM-as-a-judge evaluation, and automated grading. Such rubrics are reliable only if they reward honest answers over adversarial answers optimized to exploit them. Yet their robustness to

Emergence World: Adversarial Stress-Testing of Long-Horizon Multi-Agent Systems

👍 2

As AI agents move from bounded tasks to persistent deployments, failures can propagate through memory, tools, other agents, and environmental state long after their interactions. This creates a safety regime that cannot be characterized by evaluating model responses in isolation. Emergence World, is

Modality-Autoregressive World-Action Models

👍 7

World-action models (WAMs) jointly model future observations and actions, typically predicting the future as RGB images. Other visual modalities such as depth, pretrained visual features, and point tracks can more efficiently capture geometric, semantic, and motion features. However, how best to com

Register Tokens for Bounded-State Reasoning in Diffusion Language Models

👍 4

Masked diffusion language models (dLLMs) generate text by iteratively denoising masked tokens with bidirectional attention. Extending reasoning across generation chunks normally requires keeping earlier generated text in context. We ask whether a dLLM can instead continue reasoning after that text i

ModularRSI: Modular and Generalizable Recursive Harness Self-Improvement

👍 10

Recent work extends recursive self-improvement (RSI) to agent harnesses for long-horizon coding and terminal tasks, enabling agents to improve execution mechanisms from experience. However, generalizable harness RSI remains challenging. First, evolving harnesses on evaluation benchmarks or their sub

HarnessVLN: Unifying Training-Free Embodied Navigation through an Agent Harness

👍 17

Embodied navigation requires agents to interpret visual observations, accumulate spatial knowledge, and execute actions to follow instructions or locate objects. Training-based methods face generalization challenges, while training-free methods exploit multimodal large language models (MLLMs) but of

Decoy Direction Optimization: A Post-Hoc Defense Against LLM Abliteration

👍 2

Safety guardrails in open-weight language models can be readily bypassed using Refusal Feature Ablation (RFA), a technique that identifies and projects out a linear refusal direction from the residual stream, often achieving a high attack success rate (ASR) while preserving model capability. Defendi

Flattening Every Memory Peak in Long-Context Mixture-of-Experts Training

👍 9

Training a Mixture-of-Experts (MoE) model at long context or large batch size fails as soon as any one component's peak allocation exceeds device memory, so the target is every peak at once, not the average footprint. Four are left unbounded by the parallelism plans in common use, and each grows dif

Another Blueprint In The Wall: How to Ask Frontier AI Like a Kid?

👍 7

This paper reports experiments across six frontier model types from OpenAI, Anthropic, xAI, and Google DeepMind. Ten independent sessions per model type used the same three stage prompt sequence, progressing from architectural preference to a full ASCII backbone. Under the school audience framing, r

StepAudio 3 Realtime Technical Report

👍 102

Realtime spoken interaction demands deep reasoning, prompt responses, and fluid turn-taking. We present StepAudio 3 Realtime, an audio-language foundation model organized around a continuous listen-converse-think-act loop. Deep Perception captures rich acoustic cues to interpret user intent, while S

StepAudio 3 Music Technical Report

👍 76

We introduce StepAudio 3 Music, a large-scale, long-form music generation model that supports explicit musical planning and open-domain text-controlled generation. The StepAudio Music Tokenizer represents audio as a 50-Hz stream from a 65536-entry single codebook, using semantically informed self-su

The Last AI Built by Humans: Toward Genuine Recursive Self-Improvement

👍 88

Recursive self-improvement (RSI) enables AI systems to turn experience and feedback into persistent changes that improve both their capabilities and the process of future improvement. We first use the Headroom-Closed Index (HCI) to reveal the problems of existing LLMs, then introduce the RSI concept

Its time to talk about ‘model welfare’

@mustafasuleyman · 1.1M 粉丝 · 204.3K 阅 · 502 赞 · 86 转

AIs are not conscious. They do not feel, experience, or suffer. However, there’s a growing movement in support of model welfare; the idea that we might soon owe them a duty of care. I think this

中文介绍 探讨“模型福利”,即人工智能模型应获得关怀的观点。

THE ARCHITECTURE OF AI-NATIVE WEB3.0

@LingoAITech · 411.4K 粉丝 · 156.5K 阅 · 524 赞 · 227 转

From the Semantic Web to the Human-Centric Agentic Economy Data3.0 + Finance3.0: Sovereign Data, Personal AI, Open Agents and Programmable Value Abstract Artificial intelligence is changing the role

中文介绍 分析AI原生Web3.0架构,从语义网到以人为中心的智能经济。

open the frontier

@jack · 12.1M 粉丝 · 153.5K 阅 · 2.2K 赞 · 304 转

the frontier is the edge of what we know. no company owns what comes next. i want more people to be able to advance it. i favor open releases that people can examine, use, and improve together without

中文介绍 提倡开源AI前沿技术,鼓励共同探讨、使用和改进。

Adopting the software factory model: crawl, walk, run

@zachlloydtweets · 20.4K 粉丝 · 123.1K 阅 · 500 赞 · 47 转

The software factory approach (a closed agentic loop that runs in the cloud) is growing in popularity, but it can be daunting to adopt. In this post I’ll go through “crawl, walk, run” steps for making

中文介绍 介绍软件工厂模型采用步骤:爬行、步行、奔跑。

How we built Hermes to support our entire team

@JacquelineSYC19 · 1.0K 粉丝 · 72.4K 阅 · 501 赞 · 47 转

At the time of writing, Artie is a team of 17. Every one of us works alongside Hermes, an open-source AI agent harness from @NousResearch. It runs on a single physical machine in Germany, has a

中文介绍 分享构建支持团队工作的开源AI代理Hermes的经验。

Introducing the DeepMind Institute

@ShaneLegg · 82.3K 粉丝 · 40.6K 阅 · 585 赞 · 86 转

We are on the cusp of a profound transformation. Today’s AI systems have impressive capabilities and the rapid pace of innovation suggests we’re now approaching artificial general intelligence (AGI),

中文介绍 介绍DeepMind Institute,展望迈向通用人工智能的进程。

Build real-time voice applications with Gemini 3.8 Live and 3.5 Transcribe

@GoogleAIStudio · 203.4K 粉丝 · 35.5K 阅 · 506 赞 · 51 转

New Gemini Audio models are available for developers to build more intelligent conversational experiences via the Gemini API and Google AI Studio. Today, we released new Gemini Live models in the

中文介绍 介绍Gemini 3.8 Live和3.5 Transcribe,构建实时语音应用。

Introducing Gemini 3.8 Live and 3.8 Live Extended Thinking

@GoogleAIStudio · 203.4K 粉丝 · 35.1K 阅 · 585 赞 · 47 转

Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking are our most advanced live dialogue models yet. Major upgrades in intelligence and parallel reasoning make it more intuitive to collaborate and

中文介绍 介绍Gemini 3.8 Live和3.8 Live Extended Thinking,提升智能对话能力。

AI Agents Are Moving Beyond Chat

@Didotxyz_ · 64.6K 粉丝 · 5.7K 阅 · 502 赞 · 325 转

The important shift in AI agents is not the disappearance of chat. It is that chat is no longer the whole product. An assistant explains a task. An agent system has to carry it forward: call a tool,

中文介绍 AI代理发展超越聊天,转向任务执行和工具调用。

Introducing Astra for Law

OpenAI for Law brings frontier intelligence for law, custom firm workflows, connected legal data sources, and legal-grade controls for confidential client work.

中文介绍 OpenAI推出Astra for Law,提供法律领域的先进智能和定制化工作流程。

Our framework for reporting model misalignment

OpenAI shares a framework for tracking, investigating, and disclosing model misalignment, alongside six reports of unexpected or concerning model behavior.

中文介绍 OpenAI分享模型偏差报告框架,包括六份关于意外或令人担忧的模型行为的报告。

Helping older adults use AI in everyday life

OpenAI and AARP are bringing free, hands-on ChatGPT workshops to 1,000 older adults across 10 U.S. cities to build practical AI skills safely.

中文介绍 OpenAI与AARP合作,为1000名美国老年人提供免费的ChatGPT研讨会,以安全地构建实用AI技能。

Reimagining advertising with AI

Explore new AI-powered advertising experiences from OpenAI, including Sponsored Agents, tools for marketers, and integrations with HubSpot and Shopify.

中文介绍 OpenAI探索新的AI驱动广告体验,包括赞助代理、营销工具和与HubSpot和Shopify的集成。

Building the materials foundation for AI

The AI boom is becoming a materials challenge. As AI pushes computing into new territory, the materials behind that infrastructure are becoming just as crucial as the algorithms running on it. Semiconductors and data centers are approaching physical limits around performance, thermal management, ele

中文介绍 AI繁荣成为材料挑战,半导体和数据中心的性能接近物理极限。

How to connect AI usage to business value

Learn how ChatGPT Work and Codex analytics help teams understand AI usage and spend, identify training needs, and connect adoption to business outcomes.

中文介绍 学习如何通过ChatGPT Work和Codex分析连接AI使用与业务价值。

How workers are unlocking new ways of working

New OpenAI Economic Research shows how workers use AI beyond traditional roles and which new activities become recurring parts of their work.

中文介绍 OpenAI经济研究显示,工人如何使用AI超越传统角色,哪些新活动成为他们工作的常规部分。

Can Skills Learned in Games Transfer to Real-World Work?

Good Start Labs trained an AI on a railroad game — and one version improved at financial research. The difference was the training design.

中文介绍 Good Start Labs训练AI在铁路游戏上,一个版本在金融研究上有所改进,差异在于训练设计。

Roundtables: Could AI really kill us all?

Listen to the session or watch below Employees at the world’s leading AI labs are saying there’s a real possibility that advanced AI could destroy humanity. Are they right? Or is this more scaremongering and hype? Watch a conversation unpacking AI extinction fears: where they come from, whether they

中文介绍 世界领先AI实验室的员工表示,高级AI有可能毁灭人类。这是真的吗?还是这只是恐慌和炒作?

Australia news live: OpenAI warns Labor about global security threats; Albanese to meet Apple’s Tim Cook in US

Follow the day’s news live Get our breaking news email, free app or daily news podcast Australia needs to slow down its artificial intelligence use and create better security laws, the companies creating the technology have told the federal government. In a submission to an ongoing inquiry, OpenAI s

中文摘要 OpenAI 告诫澳大利亚政府减缓人工智能使用,并制定更严格的网络安全法律。

Yemen War Intensifies as Houthi Militia Advances on Marib

Clashes between Iranian-backed Houthi fighters and Saudi-backed government forces have been reported near oil-rich Marib Province.

中文摘要 胡塞武装在也门马里布省与沙特支持的政府军发生冲突。

YouTube star Ms Rachel to match Macklemore’s $1m Palestine aid donation

Educator talks of ‘moral obligation’ and urges ‘every wealthy white celebrity’ to make similar donation YouTube star Ms Rachel has said she is matching rapper Macklemore’s $1m donation to organizations supporting Palestinians, calling on other public figures to follow suit. Macklemore, who was fired

中文摘要 YouTube 明星 Rachel 捐赠 100 万美元支持巴勒斯坦,并呼吁其他名人效仿。

Russia election: Could other parties challenge United Russia?

Russians vote for the first time since 2022. Can opposition parties challenge Putin-backed United Russia in this vote?

中文摘要 俄罗斯首次投票以来,反对党能否挑战普京支持的统一俄罗斯党。

Trump administration approves sale of F-35 jets to Saudi Arabia

The deal, which needs approval from Congress, comes as Riyadh seeks Washington's help in its war with Yemen's Houthis.

中文摘要 特朗普政府批准向沙特阿拉伯出售 F-35 战机,以帮助其对抗也门胡塞武装。

Comedians Beyond Borders

A new generation of cross-cultural comics is taking regional humor global.

中文摘要 新一代跨文化交流喜剧家将地区幽默带向全球。

Pennsylvania seeks CDC help amid dispute over US measles deaths

The CDC and officials from the US state differ over how four deaths should be classified as cases continue to spread.

中文摘要 宾夕法尼亚州就美国麻疹死亡病例分类与疾病控制与预防中心发生争议,寻求其帮助。

Asian Stocks to Gain on Lower Oil, US Bonds Rally: Markets Wrap

Asian stocks and bonds were set to rise, tracking a Wall Street rally as a pullback in oil prices eased inflation concerns and helped revive appetite for risk. The yen was steady ahead of the Bank of Japan’s rate decision.

中文摘要 亚洲股市和债券预期上涨,受华尔街上涨推动,油价回落缓解通胀担忧,风险偏好回升。

Latest Oil Market News and Analysis for Sept. 18

Oil dropped for a third day as supply concerns eased and traders looked to the next round of diplomacy that’ll shape the US-Iran war.

中文摘要 油价连续第三天下跌,供应担忧缓解,交易员关注下一轮将影响美伊战争的外交。

OpenAI discloses new ‘concerning’ model behaviour

Developer launches system to track and report AI model misconduct

中文摘要 OpenAI披露新的人工智能模型行为问题,开发者推出系统跟踪和报告AI模型不当行为。

01.AI's Kai-Fu Lee on Pacing Frontier AI Development

Kai-Fu Lee, CEO of 01.AI, says he suspects American AI companies may already have developed models that are "really great, but also really scary." He discusses this on "Bloomberg: The China Show" on the sidelines of the BNP Paribas Global Markets APAC Conference. (Source: Bloomberg)

中文摘要 01.AI的Kai-Fu Lee表示,怀疑美国AI公司可能已开发了“非常出色但也很可怕”的模型。

Macerich CEO: Consumers Are Spending, but Are Selective

Macerich's President and CEO, Jack Hsieh, joins the program shortly after presenting at the Bank of America real estate conference. The segment is framed around Macerich and the broader theme of "tracking the shopping mall resurgence." He speaks with Romaine Bostick on "The Close." (Source: Bloomber

中文摘要 Macerich总裁兼首席执行官Jack Hsieh表示,消费者在消费,但选择很谨慎。

Schwab Is Neutral on Equities, Favorable on Commodities, Sonders Says

Liz Ann Sonders, Charles Schwab chief investment strategist, says there is no one asset allocation that makes sense for every investor. She says the firm is neutral on equities, less favorable on fixed income, and more favorable on commodities. (Source: Bloomberg)

中文摘要 Charles Schwab首席投资策略师Liz Ann Sonders表示,没有一种资产配置适合所有投资者,该公司对股票持中性态度,对商品持乐观态度。

BNY Chief Economist on Fed Rate Hike, BOJ Outlook

Vincent Reinhart, Chief Economist over at BNY Investments, discusses the Bank of Japan’s upcoming rate decision due Friday. He speaks with Romaine Bostick on "The Close." (Source: Bloomberg)

中文摘要 BNY投资首席经济学家Vincent Reinhart讨论日本央行即将到来的利率决定。

DoubleLine’s Gundlach Warns of Fiscal Crisis in Next Recession

DoubleLine Capital chief executive Jeffrey Gundlach warned that the next US downturn could trigger a debt crisis that sends long-term Treasury yields sharply higher — defying decades of conventional wisdom that bonds will always serve as safe haven during times of economic strife.

中文摘要 DoubleLine Capital首席执行官Jeffrey Gundlach警告称,下一次美国经济衰退可能引发债务危机,长期国债收益率将大幅上升。

US to Sell F-35s to Saudi Arabia in $24.3 Billion Deal

The State Department notified Congress that it plans to sell Saudi Arabia up to 48 F-35 warplanes in a first-time sale that could be worth as much as $24.3 billion. The sale of the Lockheed Martin Corp. fighter jet to Saudi Arabia has been in the works for a number of years. President Donald Trump s

中文摘要 美国国务院通知国会,计划向沙特阿拉伯出售最多48架F-35战机,交易价值可能高达243亿美元。

Jeni’s Splendid Ice Creams Founder on New Fiber Company

Jeni Britton, founder of Jeni’s Splendid Ice Creams and now Founder and CEO of Floura, discussed her shift from ice cream to fiber-based snacks made from produce trim like apple cores, watermelon rinds, mango and pineapple rinds. Britton highlighted Starbucks as an important validation point for the

中文摘要 Jeni Britton,Jeni’s Splendid Ice Creams创始人,现在为Floura的创始人和首席执行官,讨论了她从冰淇淋转向以苹果核、西瓜皮、芒果和菠萝皮等水果边角料制成的纤维零食。

Carry Trade Explained: Why the Japanese Yen Is Losing Its Appeal (JPY/USD)

For decades, the Japanese yen has been one of the world’s favored currencies for investors looking to borrow cheaply and put that money into assets offering higher returns.

中文摘要 几十年来,日元一直是投资者青睐的货币,他们希望以低息借款并将资金投入到提供更高回报的资产中。

该源今日无内容。