每日简报

2026-09-23

← 历史归档

anthropics/financial-services

Python · ★ 36,296 · 🍴 5,310 · 📈 436 stars today

中文介绍 提供金融服务的综合平台,可能用于构建金融应用程序和解决方案。

agent-substrate/substrate

Go · ★ 2,928 · 🍴 381 · 📈 301 stars today

Agent Substrate: the core system

中文介绍 Agent Substrate 是一个核心系统,可能用于构建智能合约和去中心化应用。

dream-num/univer

TypeScript · ★ 15,335 · 🍴 1,367 · 📈 202 stars today

The Office Harness for AI Agents — Spreadsheets, Docs, Slides, Canvas, Relational Tables, and PDF in one runtime.

中文介绍 为人工智能代理提供办公工具,包括电子表格、文档、幻灯片、画布、关系表和PDF,适用于多场景协作。

davila7/claude-code-templates

Python · ★ 31,084 · 🍴 3,536 · 📈 113 stars today

CLI tool for configuring and monitoring Claude Code

中文介绍 Claude Code 的配置和监控CLI工具,适用于开发者快速搭建和监控Claude Code应用。

google/ax

Go · ★ 7,492 · 🍴 350 · 📈 2,324 stars today

Google's open agentic orchestration runtime

中文介绍 Google 开源的代理编排运行时,可能用于构建自动化和智能代理系统。

mvt-project/mvt

Python · ★ 14,045 · 🍴 1,351 · 📈 441 stars today

MVT (Mobile Verification Toolkit) helps with conducting forensics of mobile devices in order to find signs of a potential compromise.

中文介绍 MVT (移动设备取证工具包) 用于对移动设备进行取证分析,查找潜在的安全威胁。

superdesigndev/treg

Python · ★ 2,173 · 🍴 215 · 📈 197 stars today

OpenRouter for agent tools. Join community here: https://discord.gg/6mQYYfFMAn

中文介绍 OpenRouter 是一个代理工具的开放路由器,适用于智能代理工具的开发和社区协作。

browser-use/video-use

Python · ★ 25,798 · 🍴 3,115 · 📈 155 stars today

Edit videos with coding agents

中文介绍 使用编码代理编辑视频,适用于视频制作和内容创作。

ACLArena: Agent Continue Learning in Multi-stage Post-training

👍 6

Building general-purpose agents for industrial deployment requires integrating multiple capabilities, each typically acquired at a distinct stage of training. Yet there is currently no well-established recipe for Agent Continual Learning (ACL), with little understanding of the trade-offs among exist

1% of Tokens Can Be Enough: On Gradient Estimation in On-Policy Distillation

👍 7

Sparse on-policy distillation (OPD) allocates teacher supervision to a small subset of tokens in student-generated trajectories. However, useful teacher guidance can yield a noisy update when its gradient is estimated from a sampled next token. We study this estimation problem at a fixed prefix in i

Harness-Zero: Harness Distillation via Agent-as-Harness

👍 12

Agent harnesses, the external systems that mediate model-environment interaction, can substantially improve agent performance, but their gains remain tied to the harness at deployment. Because the best harness varies across domains, instances, and models, a general-purpose agent must either settle f

Jev-Mem: System-One-Controlled Agentic Memory for Efficient AI Agents

👍 11

Agentic memory is becoming essential for long-horizon AI agents, yet many existing systems rely on autoregressive LLMs to control how memories are organized, retrieved, and used, placing expensive generation on the critical path of memory operations. We introduce \method, a new agentic memory archit

VideoGen-Agent: Reinforcing Video Generation Agents

👍 35

Recent advances in video generative models have enabled high-fidelity, temporally coherent video generation. However, these models often struggle to satisfy prompts requiring specialized knowledge, specific identities, physical consistency, or ordered events. In this paper, we present VideoGen-Agent

RRSI: Regularized Recursive Self-Improvement of Agent Harnesses

👍 152

An LLM agent's capability is largely magnified by its harness, namely the prompts, control flow, tooling, memory, and context management surrounding the frozen backbone model. Recent methods increasingly automate this process by iteratively proposing and selecting component-wise edits of an agent ha

Mira-Scene: Pixel-Aligned Layouts for Generative 3D Scene

👍 3

Single-image 3D object generation can now produce high-fidelity assets, yet accurately placing them into a coherent scene layout remains an open challenge. A central difficulty lies in how object layout is represented. Holistic methods absorb placement into a scene-level generation process, sacrific

Grounded Action Model: 3D Grounding as a Foundation for Robotics

👍 19

Manipulation policies must know which objects matter and where they are, yet the pretrained backbones that current robot foundation models build on, from language in vision-language-action models (VLAs) to video generation in world-action models (WAMs), do not directly require this metric grounding,

Towards Full Pipeline FP8 Reinforcement Learning for LLMs

👍 3

Reinforcement learning (RL) has become a key technique for improving the reasoning and agentic abilities of large language models (LLMs). Although FP8 quantization can accelerate RL training, maintaining stability throughout an FP8 RL pipeline remains challenging. While previous works have focused o

UltraTex: Unleashing 2K Multi-View Diffusion for 3D Texturing

👍 1

High-quality texture generation is essential for creating realistic and production-ready 3D assets. Recent multi-view diffusion methods have shown promising results for image-guided 3D texturing, but they are typically constrained to low operating resolutions such as 512 or 768, making it difficult

OmniEdu: Open Foundation Models for Learning and Teaching

👍 66

Educational foundation models must solve problems, understand curriculum structure, diagnose learner difficulties, and provide appropriate instructional support. Existing educational language models often focus on either problem solving or tutoring, with training mixtures organized by source or task

Transferring the Intelligence of VLMs to Robotic Control

👍 101

Humans can seamlessly adapt to both physical and digital worlds, suggesting that while a digital-to-real gap exists in embodiment, environment and task, human intelligence itself may transfer across this gap. This naturally raises a fundamental question: can the intelligence of vision-language model

The Functionalizer: Lossless Functional Decomposition for Subword Tokenization

👍 6

Standard subword tokenizers either treat every orthographic variation of a word (such as hello, Hello, HELLO, and Héllo) as unrelated vocabulary entries, which fragments the embedding space, or discard this variation through lossy normalization. We present the Functionalizer, a lossless pre-tokenize

Prediction-Powered Smoothing and Validation for Disaggregated AI Evaluation

👍 1

Evaluating an AI system requires disaggregated assessment, as performance varies across domains such as benchmark task types or conversation types in deployed agents. Exhaustive testing is expensive, so evaluation rests on a sample of labeled units. We treat the evaluation set as a finite population

BI-Agent and BI-Bench: Towards Automating End-to-End Business Intelligence

👍 15

Business intelligence (BI) is a cornerstone of enterprise decision-making and is widely used by enterprise users in software such as Power BI and Tableau. In traditional BI workflows, users need to prepare data by (1) identifying relevant tables, (2) performing data transformations, and (3) building

Realtime-Venus: A full-duplex interaction system with asynchronous delegation

👍 11

Natural interaction in digital and physical environments requires continuous perception and timely responses. Spoken dialogue relies on acoustic and linguistic cues, while video interaction also requires grounding the conversation in evolving visual context. We present Realtime-Venus, a proactive fu

SteerDuplex: Steerable Duplex Speech Dialogue Models

👍 2

Full-duplex spoken dialogue models support low-latency turn taking, interruption handling, and backchanneling, yet a key capability remains underexplored: steerability, the ability to reliably shift conversational behavior along attributes such as tone, persona, speaking rate, and voice style in res

SkillSpec: Intent-Masked Specification Reasoning for Agent Skill Correctness

👍 3

Autonomous agent systems increasingly depend on reusable skill abstractions for consolidating experiential knowledge and domain expertise. These artifacts typically bundle free-form instructions with heterogeneous resources. However, ensuring their correctness remains challenging. Their failure mode

ShieldVLA: Feasibility-Aware Safety Alignment for Vision-Language-Action Models

👍 2

Vision-Language-Action (VLA) models demonstrate strong generalization in robotic manipulation and navigation, but existing fine-tuning methods provide limited safety guarantees. Current approaches primarily rely on Lagrangian optimization that enforces safety through soft penalties on expected cumul

Measuring the Checker: Mutation Analysis for GPU-Kernel Benchmark Oracles

👍 3

Benchmarks for LLM-generated GPU kernels decide correctness with a few random inputs and a loose floating-point tolerance, and their verdicts now feed leaderboards and reinforcement-learning rewards. Recent work agrees these checkers are weak and patches them by hand---extra input distributions, fuz

Jev is INSANE for Marketing

@dsqjaffa · 737 粉丝 · 217.8K 阅 · 510 赞 · 37 转

You can't escape it. Literally EVERYWHERE you look on your feed is filled with something about "jEv is iNsAnE"... and it's generated tens of millions of views on X in less than a week. I'm guilty of

中文介绍 博主分析 jEv AI 模型的营销策略,短短一周内获得数百万次观看。

You're Already a Meat Proxy

@obie · 13.1K 粉丝 · 157.2K 阅 · 738 赞 · 98 转

I’m excited about a future that I suspect will be very hard on many long-time friends and certain people that I love. The same technology that is allowing me to be more successful than ever is

中文介绍 博主探讨科技进步对个人关系的影响,技术进步让某些人成功的同时,也让一些人感到压力。

I reverse-engineered Instinct's memory. Here's exactly how it works

@DhravyaShah · 62.1K 粉丝 · 120.2K 阅 · 509 赞 · 16 转

Instinct has taken the world by storm over the last two weeks - it's one of the best imessage assistants I've used. as with every product, I decided to reverse-engineer Instinct's memory to find out

中文介绍 博主逆向工程 Instinct 消息助手,分析其工作原理。

Jev + graphical models: a paradigm shift?

@fdellaert · 16.7K 粉丝 · 74.9K 阅 · 504 赞 · 61 转

Could Jev + graphical models fundamentally change how we build systems that reason under uncertainty? Unless you've been off the grid this week, you now know that Jev is @typesafeai’s new model for

中文介绍 博主探讨 Jev 和图形模型结合的可能性,可能改变系统推理的不确定性处理方式。

A Beginner's Guide to Jev

@omarsar0 · 320.6K 粉丝 · 73.5K 阅 · 523 赞 · 60 转

As you may have seen recently, Jev (a generalist System One model) took Twitter by storm. Many misleading examples are everywhere, but they don't really show Jev's real capabilities. Hence, we

中文介绍 博主介绍 Jev 的入门指南,纠正对 Jev 的误解。

Jev, System One models, and the future of computer use

@trycua · 19.0K 粉丝 · 59.5K 阅 · 501 赞 · 39 转

What we've been building with decision models at Cua - by @ddupont808 and @francedot Over the last few days we've deprioritized some other shiny objects because we've all been Jev-crazy. TypeSafe

中文介绍 博主介绍 Cua 公司利用决策模型与 Jev 的工作,探讨计算机使用未来。

Jev Explained for Normies

@matthewcanham · 10.7K 粉丝 · 57.1K 阅 · 516 赞 · 45 转

If you’ve been hanging out on the internet the past few days, you’ve probably come across Jev, a new AI model that’s taken the AI community by storm. Jev is not a traditional LLM. It’s built for

中文介绍 博主为非专业人士解释 Jev AI 模型,指出其与传统大型语言模型的不同。

Jev-as-a-Judge for Agent Evals

@LangChain · 265.9K 粉丝 · 46.6K 阅 · 504 赞 · 68 转

By Daniel Shea and Seán Roche Key takeways: Jev is a fundamentally different kind of evaluator. It returns typed answers directly instead of generating text like an LLM judge. Jev was dramatically

中文介绍 LangChain 推出 Jev 作为评估者,提供类型化答案,提高评估效率。

Build your own Jev (100% local)

@_avichawla · 77.5K 粉丝 · 37.8K 阅 · 593 赞 · 80 转

Everything you need to turn an open-source LLM into a fast, local decision engine without retraining it. It covers next-token scoring, fixed choices with probability distributions, SGLang, and a

中文介绍 博主介绍如何将开源 LLM 转化为本地决策引擎,无需重新训练。

Build a Jev Judge

@akshay_pachaar · 289.5K 粉丝 · 32.1K 阅 · 501 赞 · 59 转

We use LLMs to write answers, then call another LLM to judge them. But when the judgment is a handful of bounded decisions, do we need another round of text generation? An agent answers a customer

中文介绍 博主提出构建 Jev 评估者,解决判断中多个有限决策问题。

A Swarm of Fruit Flies, Building a Civilization On-Chain

@murmur_arc · 983 粉丝 · 11.8K 阅 · 553 赞 · 137 转

Prologue — No LLMs, Only Neurons This project is called murmur, and it is alive right now on Arc mainnet (Chain ID 5042), settling in real USDC. To date the flies have completed more than 23,000

中文介绍 博主介绍项目 murmur,使用神经元构建基于区块链的文明。

Better prompt caching for GPT-6

Learn how GPT-6 improves prompt caching with higher cache hit rates, new diagnostics, explicit breakpoints, and controls that reduce latency and costs.

中文介绍 GPT-6通过提高缓存命中率、新诊断、显式断点和减少延迟和成本的控制,改进了提示缓存。

Introducing GPT-6 Sol and Luna

Meet GPT-6 Sol and Luna, two models that bring frontier intelligence to everyday work with different balances of capability and cost.

中文介绍 OpenAI推出GPT-6 Sol和Luna模型,平衡能力和成本,提高日常工作中的智能。

Roundtables: The Deadly Failures of The Virtual Border Wall

The US has spent billions building a “virtual wall” of surveillance towers along its southern border over the past 25 years, promising they will help detect and apprehend border crossers and save lives. But a groundbreaking investigation by MIT Technology Review has documented over a thousand people

中文介绍 MIT科技评论调查了美国花费数十亿美元在南部边境建立的“虚拟墙”的失败。

Parallel cut research time and cost in half with GPT‑6 Astra

GPT‑6 Astra allowed Parallel’s agents to research and synthesize labor-market data in half the time and at half the cost vs. prior models.

中文介绍 GPT-6 Astra将Parallel的代理研究劳动市场数据的时间成本减半。

Don’t be fooled by this summer of AI hype

It’s been a busy few months for AI hype. At the end of April, Anthropic claimed that its model Claude Mythos is better at finding software vulnerabilities than most security experts. Then we had the OpenAI–Hugging Face hacking incident, after which Anthropic (proudly) and Meta (reluctantly) disclose

中文介绍 MIT科技评论指出,当前夏季的AI炒作是不可信的。

Priorities and principles for effective third party assessments

OpenAI outlines priorities and principles for rigorous, secure, and independent third-party AI safety assessments of frontier models and safeguards.

中文介绍 OpenAI概述了有效第三方AI安全评估的优先级和原则。

Higgsfield AI ships new video features in a day with GPT-6 Astra

With GPT-6 Astra, Higgsfield AI makes video ad creation easier for small businesses and brings new creative tools to market faster.

中文介绍 Higgsfield AI利用GPT-6 Astra推出新视频功能,为小企业简化视频广告制作并加快新创意工具上市。

U.N. Live Updates: As Trump Threatens Iran, Other Leaders Warn of Deep Divisions

President Trump said he must decide whether to make a deal or “annihilate” Iran, in a gathering where world leaders typically focus on declarations of global harmony.

中文摘要 特朗普在联合国大会上威胁伊朗,称可能进行军事行动,引发国际领导人的担忧。

Australia news live: ABC watchdog criticises factchecking and ‘culture of secrecy’ at Four Corners

Follow the day’s news live Get our breaking news email, free app or daily news podcast Bragg says government should be able to have ‘big debates’ about housing: ‘I don’t think that’s controversial’ Andrew Bragg, the shadow housing minister, said he believes the country should be able to have “big de

中文摘要 澳大利亚ABC监管机构批评《四角之地》节目的事实核查和“保密文化”。

Mexico Becomes Latest Country to Restrict Phones in Schools

President Claudia Sheinbaum has sharply limited the use of cellphones and screens in schools. The measures will go into effect in November.

中文摘要 墨西哥总统克劳迪娅·希恩鲍姆限制学校使用手机和屏幕,措施将于11月开始实施。

Trump Meets Middle East Leaders as Regional Wars Widen

The Iran war and the fight between the Houthis and Saudi Arabia in Yemen have forced Gulf Arab states to reckon with spiraling conflicts and to reassess U.S. security guarantees.

中文摘要 特朗普会见中东领导人,讨论地区冲突和美国家安全保证。

Iran and US hold talks on sidelines of UN summit, says Trump

Senior Iranian official indicates willingness to reopen strait of Hormuz if US takes steps to ease pressure on Tehran Iranian officials held three hours of talks with US special envoy Steve Witkoff on the sidelines of the UN general assembly, Donald Trump disclosed on Tuesday. He said there would a

中文摘要 伊朗和美国在联合国大会外举行会谈,伊朗官员表示愿意重新开放霍尔木兹海峡。

US signs ‘tremendous’ security deal with Denmark and Greenland

Trump and leaders of Greenland and Denmark have signed a deal allowing an expanded US military presence in Greenland.

中文摘要 特朗普与格陵兰岛和丹麦领导人签署“重大”安全协议,扩大美军在格陵兰岛的驻军。

Asian Stocks Set to Extend Gains as Tech Rallies: Markets Wrap

Asian stocks were poised to extend their winning streak after chipmakers drove the Nasdaq 100 to its first record since June, while falling oil prices provided a further boost amid hopes for progress toward ending the war with Iran.

中文摘要 亚洲股市预计将延续涨势,科技股上涨推动纳斯达克100指数创6月来新高,油价下跌进一步提振市场,希望伊朗战争结束取得进展。

Murat Kurum Outlines COP31 Action Plan

Murat Kurum, President, COP31; Minister of Environment, Urbanization & Climate Change, Republic of Türkiye discusses his action plan for the Antalya summit at the COP31 Business Forum New York Dialogue at Bloomberg Green New York 2026. (Source: Bloomberg)

中文摘要 土耳其环境、城市化和气候变化部长Murat Kurum在纽约Bloomberg Green 2026年对话中概述了COP31行动计划。

Fort Trump Will Come to Poland, Says President Nawrocki

Poland President Karol Nawrocki says it's just a matter of time before a permanent US military base is built in his country. Nawrocki says he hopes "Fort Trump" is completed before Trump's term ends in 2029. Nawrocki speaks exclusively to Bloomberg's Francine Lacqua on the sidelines of UN General As

中文摘要 波兰总统Nawrocki表示,在2029年特朗普任期结束前,美国将在其国建立永久军事基地“Fort Trump”。

CFTC Warns Prediction Markets to Reel in Risky ‘Mention’ Bets

Prediction market wagers involving bets on what high-profile people might say in public present heightened manipulation risk and can only be offered in limited circumstances, Commodity Futures Trading Commission staff warned in new guidance.

中文摘要 商品期货交易委员会警告,涉及对公众人物言论的预测市场投注存在高度操纵风险,只能在有限情况下提供。

Latest Oil Market News and Analysis for Sept. 23

Oil extended losses as the US flagged progress in talks with Iran to end their war, and Saudi Arabia moved to restart a key pipeline.

中文摘要 油价继续下跌,美国表示与伊朗结束战争的谈判取得进展,沙特阿拉伯采取措施重启关键管道。

Turkey Arrests Tera Chairman at Heart of Funds Scandal

Turkey arrested high-profile finance executive Emre Tezmen and four others as part of an ongoing investigation into imploded funds that have caused billions of dollars of investors’ money to evaporate.

New EU industry rules would damage UK, warns Burnham

The prime minister met the European Commission President during his visit to the United Nations General Assembly.

中文摘要 英国首相警告,新的欧盟行业规则将损害英国,在与欧洲委员会主席会晤期间表示。

Mike Bloomberg on the Role of Business in Climate Action

Michael R. Bloomberg, Founder, Bloomberg L.P. & Bloomberg Philanthropies delivers welcome remarks at the COP31 Business Forum New York Dialogue at Bloomberg Green New York 2026. (Source: Bloomberg)

中文摘要 Bloomberg L.P.和Bloomberg Philanthropies创始人Michael R. Bloomberg在纽约Bloomberg Green 2026年对话中发表欢迎致辞,讨论企业气候行动的作用。

COP31 Business Forum Chair on Climate Goals

Rifat Hisarcıklıoğlu, President, Union of Chambers and Commodity Exchanges of Türkiye (TOBB); Chair, COP31 Business Forum delivers welcome remarks at the COP31 Business Forum New York Dialogue at Bloomberg Green New York 2026. (Source: Bloomberg)

中文摘要 土耳其商会和商品交易所联合会(TOBB)主席、COP31商业论坛主席Rifat Hisarcıklıoğlu在纽约Bloomberg Green 2026年对话中发表欢迎致辞,讨论气候目标。

Apollo Caps Private Credit Fund Again After 14.7% Look to Exit

Apollo Global Management Inc. is limiting redemptions from a private credit fund for the third straight quarter, as its investors join the rush to pull cash from the $1.8 trillion direct lending market.

中文摘要 Apollo Global Management Inc.第三次限制从私募信贷基金中赎回,投资者纷纷撤资,该市场总额达1.8万亿美元。

South Korea Keeps Fuel Cheap for Holidays as Fiscal Cost Mounts

With millions of South Koreans set to hit the road this week during a three-day public holiday, the government is extending fuel-price caps, cushioning households from an energy shock that has pushed global oil prices above $100 a barrel.

中文摘要 韩国政府为应对假期期间的燃油价格上涨,继续实施燃油价格上限,以减轻家庭能源冲击。

该源今日无内容。