每日简报

2026-09-28

← 历史归档

paperclipai/paperclip

TypeScript · ★ 89,635 · 🍴 15,636 · 📈 2,527 stars today

The open-source app everyone uses to manage agents at work

中文介绍 Paperclip 是一款开源应用,用于管理工作中的智能代理,简化代理管理流程。

vectorize-io/hindsight

Python · ★ 37,121 · 🍴 4,825 · 📈 4,463 stars today

Hindsight: Agent Memory That Learns

中文介绍 Hindsight 是一款智能代理记忆学习工具,帮助代理从经验中学习。

debpalash/VoiceStudio

Python · ★ 39,877 · 🍴 4,740 · 📈 3,060 stars today

VoiceStudio is the open-source, fully-local ElevenLabs alternative — voice cloning, voice design, video dubbing, dictation, transcription & audiobook creation in 646 languages.

中文介绍 VoiceStudio 是一款开源的语音克隆和设计工具,支持多种语言的视频配音、语音识别、转录和有声书创作。

rohitg00/ai-engineering-from-scratch

Python · ★ 59,190 · 🍴 10,222 · 📈 848 stars today

Learn it. Build it. Ship it for others.

中文介绍 AI Engineering From Scratch 是一个学习人工智能工程实践的项目,旨在帮助用户从零开始构建和部署人工智能应用。

InfinityLoop1308/PipePipe

Shell · ★ 6,539 · 🍴 219 · 📈 139 stars today

An open-source Android app to let you browse YouTube and other services freely.

中文介绍 PipePipe 是一款开源的 Android 应用,允许用户自由浏览 YouTube 和其他服务。

vercel-labs/scriptc

TypeScript · ★ 5,362 · 🍴 138 · 📈 186 stars today

TypeScript-to-Native Compiler

中文介绍 scriptc 是一个 TypeScript 到本地编译器的工具,用于将 TypeScript 代码编译成原生应用。

mvschwarz/openrig

TypeScript · ★ 900 · 🍴 97 · 📈 114 stars today

Multi-agent harness that runs Claude Code and Codex together as one system

中文介绍 openrig 是一个多智能代理工具,可以将 Claude Code 和 Codex 结合运行,形成一个统一的系统。

dream-num/univer

TypeScript · ★ 20,113 · 🍴 1,703 · 📈 920 stars today

The Office Harness for AI Agents — Spreadsheets, Docs, Slides, Canvas, Relational Tables, and PDF in one runtime.

中文介绍 Univer 是一个适用于人工智能代理的办公套件,支持电子表格、文档、幻灯片、画布、关系表和 PDF 等功能。

willfaust/Madeira

C · ★ 786 · 🍴 152 · 📈 117 stars today

Run x86-64 Windows PC games on jailed iOS via FEX-Emu + Wine + DXMT

中文介绍 Madeira 是一个运行在 iOS 设备上的 x86-64 Windows PC 游戏模拟器,利用 FEX-Emu、Wine 和 DXMT 技术。

RGBD20K: A Large-Scale Benchmark for RGB-D Semantic Segmentation

👍 8

In this paper, we propose RGBD20K, a novel dataset for facilitating the development of more robust and general RGB-D semantic segmentation by encompassing abundant categories and high-quality annotations. RGBD20K possesses several attractive properties: (1) Expanded Semantic Space. In particular, it

Coding Agents for Generalized Task and Motion Planning Problems

👍 11

Task and motion planning (TAMP) problems remain difficult even with full observability and object-centric states because discrete decisions are tightly coupled to geometric, kinematic, and dynamic constraints. Generalized TAMP addresses this difficulty by exploiting regularities across problem insta

Rufus-Air: An Open LLM Post-Training Recipe

👍 17

Rufus-Air is an open and reproducible post-training recipe on GLM-4.5-Air-Base (106B-A12B), organized as a serial pipeline of eight stages: SFT, Reasoning RL, Coding RL, Instruction-Following RL, General Agent, Coding Agent, Search Agent, and RLHF. We document the data, reward design, infrastructure

IterSynth: Rethinking Deep Search Agents via Role-Decoupled Iterative Synthesis

👍 9

Deep search requires LLM agents to decompose complex queries, search for evidence, and synthesize grounded answers, yet existing ReAct-style agents suffer from two limitations: role coupling, where one policy must handle planning, evidence use, and synthesis; and context accumulation, where growing

PUBG Ally: A Conversational Embodied Agent as an AI Teammate

👍 6

We introduce PUBG Ally, an embodied agent for PUBG: BATTLEGROUNDS that can reason, act autonomously, and play alongside players as a voice-enabled teammate. Building such a teammate requires combining two difficult capabilities: it must perceive and respond to a constantly changing game world under

Learning to Discover Interesting Mathematics

👍 11

Recently, Large Language Models (LLMs) have been increasingly able to solve advanced mathematical problems, including many that have been open for decades. This opens the door to expansion of mathematical knowledge at unprecedented scale. Yet, while LLMs may be able to conjecture and prove more and

DeltaWAM: Delta World Action Models for Bimanual Manipulation

👍 4

World-action models (WAMs) transfer visual and motion priors from pretrained video generators to robot control by jointly modeling visual dynamics and actions. Existing WAMs, however, predict dense future frames during training, repeatedly modeling largely unchanged content and coupling action-condi

Training Object Permanence in World Models

👍 203

Object permanence and solidity are hallmarks of human cognitive priors. Recent studies show that video generation models, a paradigmatic class of current world models, have begun to show emerged reasoning abilities, making them ideal candidates for building human-like physical intelligence. Do video

Agent-Editing World Model: Rethinking World Modeling for LLM Agents

👍 19

Recent advances in large language models (LLMs) have enabled agents to tackle long-horizon tasks across diverse environments. To further improve agent performance, existing language world models typically predict environment observations, yet reconstructing high-entropy, execution-dependent tool res

FLEET: From Logits Entropy to Enhanced Trajectories in Text Generation

👍 4

Solutions based on large language models (LLMs) often rely on temperature sampling to improve accuracy and stability by aggregating multiple samples from the completion distribution. However, this memoryless approach is inherently suboptimal: because it lacks awareness of prior generations and their

Just-in-Time Memory: Learning to Curate Task-Adaptive Memory for LLM Agents

👍 35

Agentic memory systems reuse past experience to improve future performance, yet most existing designs curate memory at write time: once a task is completed, its trajectory is distilled into a fixed artifact, such as a reflection, workflow, skill, or reasoning strategy, that is later retrieved by sim

EmbodiedSWE: Coding Agents for Long Horizon Dexterous Robotics

👍 3

We study coding agents for long-horizon, dexterous robotics and ask whether their solutions can provide scalable supervision for learning general robot policies. To test this, we develop EMBODIEDSWE-BENCH, a simulation benchmark for coding agents spanning contact-rich manipulation, deformable object

StudentBench: AI and human tutoring yield equivalent GRE learning gains

👍 3

Artificial intelligence offers an unprecedented opportunity to augment human capabilities, yet progress at the frontier has focused primarily on advancing model capabilities. We introduce StudentBench, a suite of AI teaching evaluations and a public platform that enables large-scale data collection

Knowledge Pull Requests for Continual Document Authoring

👍 2

We introduce Knowledge Pull Requests (KPRs), a framework for continual document authoring that makes each change interpretable. Documents require ongoing revision as new knowledge surfaces from other sources, languages, or times, but existing approaches either edit with no account of what knowledge

Calibration as a First-Class Criterion in LLM Evaluation

👍 2

Calibration of language models -- the alignment between expressed or implicit confidence and empirical correctness -- is a well-studied subfield within NLP. Methods to measure it already exist. The problem is adoption: outside this subfield, NLP research regularly introduces new models, datasets, an

MemoryAthena: Adaptive Routing over Latent and Generated Memories

👍 5

Learned-memory methods store information in an explicit table and consume it through a separate reader, allowing addressing, storage, and reading to be modified independently. We study whether useful memory can also be generated rather than only retrieved. MemoryAthena uses three pathways: direct En

X-Planner: Event-Structured Task Planning for Embodied Intelligence

👍 2

Task planning bridges high-level instructions and executable behavior in long-horizon manipulation, yet modern Vision-Language-Action (VLA) systems often leave this intermediate structure implicit. Existing chain-of-thought (CoT) planners also tend to rely on coarse task-level annotations or seriali

HappyWorld-Bench

👍 43

Evaluating world models requires assessing both the quality of the worlds they generate and their consistency and responsiveness under exploration, interaction, and modification. We introduce HappyWorld-Bench, a comprehensive benchmark that evaluates whether generated worlds remain reliable as agent

OmniEcho: Spatial Audio Understanding for Embodied Agents

👍 22

Humans can effortlessly localize the direction of a sound source and integrate it with visual cues for reasoning, yet this remains challenging for embodied agents. In particular, it is still unclear how to effectively evaluate and model spatial audio understanding in embodied settings. To address th

Self-Organizing Agent Teams Learn to Reason Together

👍 6

Collective intelligence depends not only on what team members know, but also on how they organize their work. When the structure of a solution is unknown, useful roles and divisions of labor cannot be specified in advance; teams must learn from experience how to organize reasoning as it unfolds. Hum

AgentKernel: The Trust-Native Agentic Operating System

👍 7

Modern AI agents routinely cross trust boundaries: they ingest untrusted content, combine it with privileged instructions, persist intermediate beliefs in long-term memory, and invoke privileged tools. This creates an attack surface in which malicious payloads can enter through model inputs and caus

The New Copilot

@jacobandreou · 7.8K 粉丝 · 521.0K 阅 · 500 赞 · 77 转

The bottleneck is no longer intelligence. Over the last six months, intelligence has accelerated dramatically. But smarter models are not enough for everyone to benefit from novel intelligence. New

中文介绍 探讨AI智能瓶颈,强调模型智能提升对大众受益的重要性。

We're open sourcing the company brain, Here's how we designed the multi-player harness

@DhravyaShah · 63.2K 粉丝 · 112.4K 阅 · 608 赞 · 45 转

It's been just 2 weeks since we discontinued the company brain harness. We had to offboard tons of people to other products, but many of our customers, and others suggested us to open source our work

中文介绍 分享开源公司脑力工具设计过程,回顾产品转型过程。

Why AI is booming, but productivity isn't

@chamath · 2.4M 粉丝 · 97.8K 阅 · 551 赞 · 44 转

In June, I wrote that vibe coding was dead and that ROI-driven analysis of AI was about to go from a nice-to-have to a necessity. Here's what AI ROI is and why it's tricky to see in GDP numbers so

中文介绍 分析AI繁荣与生产力提升的差距,探讨AI投资回报率在GDP中的体现。

Prescriptions for Prosperity in the Digital Economy

@saylor · 5.2M 粉丝 · 96.2K 阅 · 661 赞 · 98 转

Artificial intelligence will make it possible for individuals and companies to produce far more than they can today. That makes the freedom to create, finance, own, and exchange things more important.

中文介绍 展望数字经济发展,强调创造、融资、拥有和交换的自由重要性。

How to build motion design studio with Opus 5.5 ( Full-course )

@0xMovez · 36.4K 粉丝 · 78.9K 阅 · 737 赞 · 46 转

Most people who try motion design with Opus 5.5 end up with the same video: centered text on a gradient, everything fading in, a logo at the end. They don't give it a reference, don't give it a

中文介绍 分享使用Opus 5.5构建动态设计工作室的教程,避免常见设计陷阱。

Using Claude Code: Spending your effort

@trq212 · 354.7K 粉丝 · 62.7K 阅 · 965 赞 · 52 转

One of the best parts of our newest Claude models is how they respond to effort without breaking the prompt cache in Claude Code, but I’ve received a lot of questions on this from users. What is

中文介绍 解析Claude Code中模型对用户努力的反应,解答用户疑问。

We built a 20x faster Grok bot

@cerebras · 76.6K 粉丝 · 37.6K 阅 · 554 赞 · 28 转

Written by @milksandmatcha and @0xSero A personal assistant should save you time and effort. Over the past few weeks, we’ve been obsessively testing AI personal assistants on everyday tasks, from

中文介绍 介绍20倍速度的Grok机器人,测试AI个人助理在日常工作中的表现。

OpenRouter: from Seed to Stripe — with OpenRouter’s Alex Atallah & AMP’s Anjney Midha

In 2023 most people doubted that there could be more than 1 or 2 frontier model labs. Now there are dozens.... and Stripe just bought the best known one for $7B.

中文介绍 Stripe以70亿美元收购了前沿模型实验室中最为知名的一家,而2023年大多数人还怀疑是否可能存在超过1或2家这样的实验室。

Proaction boosts sales 60% and saves 75+ hours with Codex

With Codex, GPT-Live-1, and GPT-6 Astra, Proaction builds, operates, and sells modern fleet management faster.

中文介绍 Proaction利用Codex、GPT-Live-1和GPT-6 Astra,使现代车队管理更快地建设、运营和销售,从而提升了60%的销售额并节省了75小时以上。

The Pentagon wants $30 million to build an AI-powered lie detector

The US government wants to spend $30.3 million over the next five years on an improved form of lie detector, according to a Department of Defense budget request. The program, called Polygraph+ or Polygraph Next, will focus on scoring algorithms that use artificial intelligence and machine learning a

中文介绍 美国国防部请求在未来五年内花费3030万美元研发一种改进型测谎仪,名为Polygraph+或Polygraph Next,将专注于使用人工智能的评分算法。

[AINews] The Future of Latent Space

A quiet day lets us discuss the work behind the scenes - now open for business!

中文介绍 Latent Space透露了幕后的工作现在对外开放。

Runway’s WorldPrompt and the Engineering of Real-Time Worlds

GWM Worlds 2 uses persistent context and timed actions to steer a world model generating video and audio in real time.

中文介绍 GWM Worlds 2利用持续上下文和定时动作来引导生成视频和音频的世界模型,实现实时世界工程。

Foundries vs Navigators: Lowering the Cost of Science

Guest Post: In science, thinking has gotten cheap but doing has not. This asymmetry is reshaping how research companies operate, largely inconspicuously.

中文介绍 科学领域思考成本降低,但实际操作成本未减,这种不对称性正在改变研究公司的运营方式。

Two years of OpenAI Academy

Marking two years of OpenAI Academy and bringing AI skills to even more communities.

中文介绍 OpenAI庆祝OpenAI学院成立两周年,致力于将AI技能带到更多社区。

🔬Bio-security is an AI Arms Race - Eric Nguyen (CEO, Radical Numerics)

Radical Numerics is using biological chain-of-thought and multimodal perception to keep up with the bio-defense arms race, design new genomes and gain insights into biology itself.

中文介绍 Radical Numerics利用生物思维链和多模态感知来跟上生物防御军备竞赛,设计新的基因组并深入了解生物学本身。

OpenAI extends cyber access to Ukraine for civilian defense

OpenAI is extending access to its Daybreak program to the Government of Ukraine to support the cyber defense of civilian infrastructure.

中文介绍 OpenAI将其Daybreak项目扩展到乌克兰政府,以支持民用基础设施的网络安全。

Sam Altman’s remarks at the United Nations Security Council

OpenAI CEO Sam Altman discusses AI safety, human control, and international cooperation in remarks to the United Nations Security Council.

中文介绍 OpenAI首席执行官Sam Altman在联合国安理会就AI安全、人类控制和国际合作发表讲话。

Australia news live: Chalmers and Gallagher to present final budget outcome; RBA interest rate hike looms

Follow the day’s news live Get our breaking news email, free app or daily news podcast Gallagher says final budget outcome will show $6bn improvement Katy Gallagher and the treasurer, Jim Chalmers, are set to present the final budget outcome today, which she said will show about a $6bn improvement s

中文摘要 澳大利亚财政部长吉姆·查尔默斯和凯特·加拉格尔将公布最终预算结果,预计将显示约60亿澳元的改善。

Serbia’s Longtime President Aleksandar Vucic Resigns Post

Aleksandar Vucic, one of Europe’s longest-serving leaders, resigned as president, paving the way for him to become prime minister after elections next month.

中文摘要 塞尔维亚长期领导人亚历山大·武契奇辞职,为下月成为总理铺平道路。

Five arrested near UK air base used by U.S. bombers in Iran war

U.K. police say five men were arrested on suspicion of preparation of a terrorist act after three vehicles were spotted near an air base the U.S. military has been using to launch attacks on Iran.

中文摘要 五名男子在英国一个被美军用作对伊朗发动空袭的空军基地附近被捕,涉嫌准备恐怖主义行动。

Floods inundate roads in southeastern Algeria

Circulating footage showed significant flooding, following heavy rain in southeastern Algeria.

中文摘要 阿尔及利亚东南部发生严重洪水,道路被淹。

Deadly strike hits market in Yemen’s Taiz

A strike on a market in the Yemeni city of Taiz has killed at least seven people and wounded 40.

中文摘要 也门塔伊兹市的一个市场遭到空袭,造成至少7人死亡,40人受伤。

A Growing North Korea Problem

In an interview with The New York Times, South Korea’s president dropped a bold proposal for how to curb North Korea’s nuclear program.

中文摘要 韩国总统提出一项大胆计划,以遏制朝鲜的核计划。

The Bond Market Is Getting Closer to Sounding Alarm on Economy

The bond market is on the brink of signaling that a series of Federal Reserve interest-rate hikes will start shifting the narrative toward the risk that the US economy stalls out.

中文摘要 债券市场接近发出经济停滞警报,预计美联储连续加息将转向风险。

Corporate America embraces cheaper ‘open’ AI models

US businesses far beyond Silicon Valley are adopting Chinese alternatives to OpenAI and Anthropic’s systems

中文摘要 美国企业广泛采用中国AI模型替代OpenAI和Anthropic的系统。

The real lesson from the Man City affair

Owners should be free to spend their own money on football — as long as they’re fit and proper

中文摘要 所有者应有权自由支配自己的足球资金,只要他们合格且适当。

Bloomberg This Weekend 09/27/2026

The news doesn’t stop when markets close. Hosts David Gura, Christina Ruffini and Lisa Mateo bring clarity, context and a bit of humor to the weekend’s biggest headlines, LIVE from New York. Joined by IAEA Director General Rafael Grossi, The New York Times Finance Reporter Maureen Farrell, Cuban Dep

中文摘要 纽约时间周末新闻节目中,主持人讨论了市场关闭后的新闻,包括IAEA总干事和纽约时报财经记者。

Pay to play in the age of corporate migration

States are feeling more pressure from companies demanding subsidies and tax breaks

中文摘要 各州面临来自要求补贴和税收优惠的公司的更大压力。

Paramount Warner Deal Tests Hollywood’s Future

Bloomberg entertainment reporter Lucas Shaw is on Bloomberg This Weekend examining Paramount Skydance’s acquisition of Warner Bros. Discovery and the challenge of turning two legacy media companies into a growing business as cable declines and streaming growth slows. Speaking with Lisa Mateo, Shaw s

中文摘要 帕拉蒙天空舞影业收购华纳探索公司,面临将两家传统媒体公司转型为增长业务的挑战。

Kennedy Center Musicians Find New Stages

Bloomberg News Washington breaking news Editor Bryan Pietsch is on Bloomberg This Weekend discussing uncertainty surrounding the Kennedy Center’s renovation, possible demolition and the displacement of major performing arts groups. He tells hosts David Gura and Christina Ruffini that musicians and o

中文摘要 肯尼迪中心音乐家在讨论翻新、可能的拆除和主要表演艺术团体搬迁的不确定性。

US Iran Talks Hit Familiar Sticking Points

Bloomberg White House correspondent and national security editor Michelle Jamrisko is on Bloomberg This Weekend discussing renewed US-Iran diplomacy after indirect talks in New York failed to overcome disagreements over sanctions, the US blockade and the Strait of Hormuz. Speaking with hosts David G

中文摘要 美国-伊朗会谈在纽约间接谈判失败后,由于制裁、美国封锁和霍尔木兹海峡的分歧,讨论了恢复美伊外交。

Tremors From AI to Oil Boost Popular Hedge Fund Dispersion Trade

Wild gyrations in individual stocks on a punchy cocktail of AI euphoria and fear, a tumbling bond market and geopolitical drama bode well for a long-favored trade among hedge fund managers.

中文摘要 人工智能和石油的波动为对冲基金经理中流行的分散交易带来利好。

Inflation Keeps Pressure on the Fed

Renaissance Macro Research economist Neil Dutta tells Bloomberg This Weekend that the US labor market has stabilized while persistent inflation could force the Federal Reserve to raise interest rates at a faster pace than investors currently expect. Speaking with hosts David Gura and Christina Ruffi

中文摘要 美国劳动力市场稳定,持续的通胀可能迫使美联储以比投资者目前预期的更快的速度提高利率。

Data Center Boom Faces Wall Street Skepticism

New York Times finance reporter Maureen Farrell tells Bloomberg This Weekend that Wall Street is growing more skeptical of the AI data center boom as rising interest rates, local opposition and permitting delays add uncertainty and costs to projects. Speaking with hosts David Gura and Christina Ruff

中文摘要 华尔街对AI数据中心热潮日益持怀疑态度,随着利率上升、当地反对和许可延误,项目的不确定性和成本增加。

该源今日无内容。