每日简报

2026-10-08

← 历史归档

morluto/rea

TypeScript · ★ 14,580 · 🍴 1,534 · 📈 4,666 stars today

Reverse engineer anything with agents, from app behavior down to native binaries.

中文介绍 通过智能代理逆向工程任何应用程序,从应用行为到原生二进制文件,适用于需要进行逆向工程的技术人员。

mattpocock/skills

Shell · ★ 279,487 · 🍴 23,412 · 📈 1,406 stars today

Skills for Real Engineers. Straight from my .agents directory.

中文介绍 为真实工程师提供的技能集合,适用于从 .agents 目录中直接使用的编程代理。

boykopovar/AnyPS5

C++ · ★ 10,228 · 🍴 778 · 📈 2,725 stars today

Tool for automatic PS5 executables porting to Linux and Windows

中文介绍 自动将PS5可执行文件移植到Linux和Windows的工具,适用于需要跨平台移植游戏或应用的开发者。

ayghri/i-have-adhd

Python · ★ 55,048 · 🍴 3,156 · 📈 620 stars today

A skill to stop your coding agent from burying the answer. ADHD-friendly output.

中文介绍 一种防止编码代理埋没答案的技能,适用于有注意力缺陷多动障碍(ADHD)的编程人员。

cathrynlavery/diagram-design

HTML · ★ 44,898 · 🍴 2,883 · 📈 828 stars today

Editorial diagram design for Claude Code, Codex, GitHub Copilot, Factory Droid, and Pi. 42 diagram types. Self-contained HTML + SVG. No shadows. No Mermaid slop.

中文介绍 Claude Code、Codex、GitHub Copilot等AI编程工具的图表设计编辑器,提供42种图表类型,支持自包含的HTML和SVG格式。

addyosmani/agent-skills

JavaScript · ★ 102,738 · 🍴 10,764 · 📈 693 stars today

Production-grade engineering skills for AI coding agents.

中文介绍 为AI编码代理提供生产级别的工程技能,适用于需要高级编程辅助的开发者。

EpicGames/raddebugger

C · ★ 7,837 · 🍴 381 · 📈 82 stars today

A native, user-mode, multi-process, graphical debugger.

中文介绍 一款原生、用户模式、多进程、图形化的调试器,适用于需要进行复杂调试的程序员。

thedotmack/claude-mem

TypeScript · ★ 97,666 · 🍴 8,599 · 📈 578 stars today

Persistent Context Across Sessions for Every Agent – Captures everything your agent does during sessions, compresses it with AI, and injects relevant context back into future sessions. Works with Claude Code, OpenClaw, Codex, Gemini, Hermes, Copilot, OpenCode + More

中文介绍 为每个代理提供会话间的持久上下文,记录代理在会话中的所有操作,并压缩后注入到未来的会话中,适用于Clau等AI编程代理。

manaflow-ai/cmux

Swift · ★ 27,815 · 🍴 2,460 · 📈 96 stars today

Open source Ghostty-based macOS terminal with vertical tabs and notifications for AI coding agents. Built for multitasking, organization, and programmability.

中文介绍 基于Ghostty的开源macOS终端,具有垂直标签和通知功能,专为AI编码代理的多任务处理和组织设计。

trycua/cua

Rust · ★ 28,710 · 🍴 2,038 · 📈 229 stars today

Scale computer-use 2.0 with open-source drivers, cross-OS fleets, and benchmarks for training, evaluation, and data generation.

中文介绍 通过开源驱动、跨平台集群和基准测试,扩展计算机使用体验,适用于需要训练、评估和数据生成的计算机使用研究者。

cloudflare/security-audit-skill

JavaScript · ★ 25,983 · 🍴 1,563 · 📈 617 stars today

A coding-agent skill for multi-phase security audits with independently verified, machine-readable findings

中文介绍 一种用于多阶段安全审计的编码代理技能,提供经过独立验证、机器可读的发现结果,适用于安全审计人员。

tester-army/e2e

TypeScript · ★ 7,344 · 🍴 331 · 📈 1,391 stars today

Next generation e2e testing framework for web and mobile apps.

中文介绍 适用于Web和移动应用的下一代端到端测试框架,适用于需要进行自动化测试的开发者和测试人员。

DuarteSantos8/openGym

JavaScript · ★ 6,805 · 🍴 897 · 📈 1,494 stars today

Self-hosted gym & body-weight tracker — plan routines, log workouts (supersets, warm-ups, cardio), see which muscles are trained, fatigued or detrained, import from FitNotes/Strong/Hevy, passkey login. Your data, your server.

中文介绍 自托管的健身与体重追踪器,可规划训练计划、记录锻炼、查看训练效果,支持从FitNotes/Strong/Hevy导入数据,使用密钥登录。

CheckerBench: Can Long-Horizon Agents Synthesize Static-Analysis Checkers?

👍 52

Static-analysis checker synthesis requires agents to interpret a defect specification, inspect a repository, implement analyzer-specific logic, and refine the checker through repeated compilation and analysis feedback. Existing coding-agent benchmarks focus on tasks such as patch generation or vulne

Sherpa: Teaching LLMs to Teach Adaptively

👍 6

Large language models (LLMs) have become increasingly capable problem solvers, but being able to solve a problem is not the same as being able to teach it. Existing approaches to training LLMs as teachers rely on demonstrations, preference data, or predefined pedagogical criteria that specify what g

DAEDALUS: Bootstrapping Agent Memory from Self-Generated Tasks

👍 9

LLM agents often lack the operational knowledge to act reliably in new environments, as they must discover specific tool behaviors or environment conventions on their own. Without memory of past attempts, they repeat the same mistakes across tasks, leading to more task failures and longer trajectori

UNREAL: Unifying Retrieval and Long-Context with a Single Model

👍 19

Long-context inference and Retrieval-Augmented Generation (RAG) handle evidence selection at vastly different scales, from a single long prompt to an entire corpus. We ask whether a single model-internal mechanism can select evidence across this range. We introduce UNifying REtrieval And Long-Contex

Towards In-Parameter Memory Augmentation for Large Language Models

👍 5

Recently Large Language Models (LLMs) and LLM-based agents increasingly need to incorporate knowledge acquired after pretraining, e.g., domain facts, user preferences, documents, and interaction experience. In-context learning (ICL) and ICL-based agent harness remain flexible, but they consume conte

From Evidence to Action: How Tool-Using Agents Fail

👍 34

Tool-using agents make consequential changes to external state, yet correct outcomes do not guarantee that their actions were supported by evidence established beforehand. We study where this evidence-to-action chain breaks as agents move from deciding whether to act to executing single actions and

Stepped MoE: Segment-Level Routing with Configurable Inference Complexity

👍 2

Training large language models (LLMs) is resource-intensive, and adapting them for diverse deployment scenarios with varying computational constraints remains challenging. While elastic architectures enable flexible model deployment and sparsely activated models allow input-adaptive computation, exi

Conditional Trajectory Peaks: Single-Pass Multimodal Policies over Action Chunks

👍 1

Multimodal imitation learning requires diverse executable futures under the same observation and consistent behavior across replanning cycles. We present Conditional Trajectory Peaks (CTP), a single-pass policy framework that jointly predicts complete action-chunk candidates, probability masses, and

Learning to Read the Contextual Tokens in Diffusion Transformers

👍 6

Multimodal Diffusion Transformers (MM-DiTs) jointly process visual and textual representations throughout generation. These models repeatedly update the text tokens through multimodal attention, forming dynamic contextual tokens whose function is not well understood. In this work, we introduce a fra

HLA: Expressive Hybrid Linear Attention via Chunk-Wise Dynamic Mixing

👍 5

Linear attention enables efficient long-context autoregressive decoding by compressing history into recurrent states, but this compression can make selective access to sparse and distant information difficult. Existing chunk-based extensions increase memory capacity, yet learned chunk-mixing coeffic

MiniCorp: The Last Mile of the AI Agent Firm

👍 15

The last mile toward enterprise AGI is a company that runs itself. Training and adapting such agents require longitudinal enterprise data, which remain scarce, costly to acquire, and often restricted by privacy constraints. Historical archives are also frequently incomplete and record only what actu

DiffGate: Difficulty-Gated Teacher Guidance for On-Policy Distillation

👍 12

On-policy distillation (OPD) has emerged as a widely used paradigm for post-training large language models, reducing the train--test mismatch of conventional distillation by supervising the student on its own generated trajectories. However, existing OPD objectives remain largely token-local and out

Magic-W0: A Structured World-Action Foundation Model for Physical Intelligence

👍 1

World-action models (WAMs) augment robot policies with action-conditioned environment dynamics, yet existing approaches largely rely on future observation reconstruction or generic latent prediction and lack structured, control-oriented world representations tightly coupled with action generation. W

ConEx: Human-Interpretable Saliency Maps via Concept-Aware Attribution

👍 1

Many visual explanation methods in computer vision highlight pixel importance but struggle to link these low-level cues to semantically meaningful concepts, limiting their interpretability and trustworthiness. We introduce Concept-based Explanations (ConEx), a novel framework that bridges saliency v

Cross-Lingual Alignment for Decoder-Only Models using MoE Routers

👍 1

Cross-lingual contrastive learning has been a core component of multilingual encoder training, but the ability to explicitly align representations is not possible in decoder-only LLMs because of varying multilingual tokenization. However, a growing amount of research suggests that even in LLMs, high

Harness-Aware Distillation for Small Language Model Agents

👍 5

Language model agents are deployed with a harness, the software around the model that manages its context, tools, and feedback. When such an agent is distilled into a smaller one, the harness stays in place, so the student mainly needs the teacher-specific abilities that the harness cannot provide,

Learning Functional Subspaces for Neural Network Compression

👍 3

Modern transformers pair impressive capabilities with substantial memory and compute demands. Low-rank weight factorization reduces both while keeping the matrices dense, and thus efficient on standard hardware. Existing methods, however, choose the subspace to remove from each weight matrix with lo

Taming VLAs under Robot Execution Errors: Self-Compensation and Stress Testing

👍 24

Vision-language-action (VLA) policies often fail when a robot's executed motion deviates from their commanded action. Such execution errors arise from the robot's mechanics and operating conditions, such as wear and payload changes. We propose self-compensating VLA, a deployment-time adaptation meth

Multilinguality in Hybrid Attention LLMs

👍 3

In response to the growing demand for long sequences in agentic and reasoning use cases, many state-of-the-art LLMs combine multiple variants of attention to mitigate the quadratic complexity of traditional softmax attention. These hybrid attention LLMs aim to balance the strengths and limitations o

Adaptive Latent Capacity for World Models

👍 10

We introduce Adaptive LeWorldModel (ALeWM), a world model based on a joint-embedding predictive architecture (JEPA) that learns to concentrate predictive information in compact prefixes of a wide latent representation. To encourage this ordering, ALeWM learns a sequence-conditioned distribution over

He Who Controls the Compute Controls the Universe

@stevenmarkryan · 458.8K 粉丝 · 124.9K 阅 · 541 赞 · 84 转

Incredible things are happening in AI. What’s more incredible is that it’s still EXTREMELY early. As in, Day 0.000. Things are only going to get crazier from here. In case you’ve been living under a

中文介绍 博主探讨AI发展现状,认为目前仍处于非常初级的阶段,未来将更加疯狂。

Every language learner gets a tutor now

@TryLiveAvatar · 1.1K 粉丝 · 123.4K 阅 · 692 赞 · 188 转

The best way to learn a language was always a patient tutor. For the first time, everyone can have one. Here's what that changes for learners, teachers and schools. The hardest part is speaking Ask

中文介绍 博主介绍语言学习AI助手,强调其对学生、教师和学校带来的变革。

AI won't take your job anytime soon. In 10 yrs, only 5% of what humans do will be replaced by AI

@mustafasuleyman · 1.2M 粉丝 · 99.2K 阅 · 587 赞 · 104 转

This is the prediction Nobel laureate Daron Acemoglu makes in the first issue of The Humanist Review, our new magazine exploring the future of AI, published by MAI. He argues we need to stop

中文介绍 博主分享诺贝尔奖得主Daron Acemoglu的预测,未来10年内AI将取代人类工作的比例仅为5%。

If you have too many interests, build this AI system in 1 hour

@Hesamation · 90.6K 粉丝 · 21.4K 阅 · 618 赞 · 51 转

if you have an ADHD brain, you probably have six different goals, you get overwhelmed, and you miss days without working on any one of them. that's me. I love agents, harnesses, RL, LLM architecture,

中文介绍 博主分享针对多兴趣人士的AI系统构建方法,利用代理、 harnesses、RL和LLM架构实现高效目标管理。

Helping teens learn, plan, and shape the future of AI

College Planner is coming to ChatGPT for Teens to help students manage college applications, alongside new flashcards, quizzes, and a teen AI council.

中文介绍 OpenAI将为青少年推出College Planner插件,帮助他们管理大学申请。

Radisson Hotel Group brings hotel discovery into ChatGPT

Radisson partnered with Accenture to build a ChatGPT plugin using OpenAI technology, helping travelers find, compare, and book hotels while planning their trips.

中文介绍 Radisson酒店集团与Accenture合作,在ChatGPT中推出酒店发现插件。

GPT-6 and Intelligent UI for everyone

GPT‑6 is rolling out globally in ChatGPT with Intelligent UI, delivering faster responses with visuals and interactive experiences you can explore and use directly.

中文介绍 OpenAI在全球范围内推出GPT-6,配备智能UI,提供更快的视觉交互体验。

How Jump Trading is scaling quant research with ChatGPT

Jump Trading uses OpenAI to expand quantitative research. See how longer-running AI workflows combine multiple data sources with human review.

中文介绍 Jump Trading使用OpenAI扩展量化研究,结合多数据源和人工审核。

Sharing AI progress in mathematics

OpenAI publishes new results on open problems in mathematics from an internal frontier model and shares Lean proof formalizations and research details on GitHub.

中文介绍 OpenAI发布关于数学开放问题的内部前沿模型新结果。

Advancing computer use with Ironclad

Learn how OpenAI and Ironclad are training and evaluating AI agents on complex contracting workflows to advance computer use for professional work.

中文介绍 OpenAI与Ironclad合作,训练和评估AI代理在复杂合同流程中的应用。

Connecting AI agents to enterprise knowledge

For all the data that AI systems continually amass and analyze, enterprise AI agents often suffer from a curious shortcoming: a lack of knowledge. More than data, knowledge is the understanding of what the data means in the context of individual organizations. AI agents need this understanding to re

中文介绍 企业AI代理缺乏知识,需要理解数据在组织中的意义。

What Life in Gaza Is Like Now

Nearly a year after Israel and Hamas agreed to end the war in Gaza, Saher Alghorra, a photographer for The New York Times, documents how residents are coping and working to rebuild their lives.

中文摘要 加沙地带战争结束近一年后,纽约时报摄影师Saher Alghorra记录了居民如何应对并重建生活。

The School Near Paris That Shows Why French Students Are Protesting

Angered by poor classroom conditions, students at Paul Éluard High School in a suburb of the capital were among the first to blockade their campus during an ongoing round of unrest.

中文摘要 巴黎近郊Paul Éluard高中的学生因教室条件恶劣,在持续的不满中首先封锁校园。

Trump on why he thinks he deserves the Nobel Peace Prize

US President Donald Trump said it would be a ‘great discredit’ to the Nobel Peace Prize committee.

中文摘要 美国总统特朗普认为他应获得诺贝尔和平奖,否则是对诺贝尔和平奖委员会的“极大羞辱”。

‘Maricarmen,’ Whose Eviction Enraged Spaniards, Dies at 87

María del Carmen Abascal, aged and ailing, became a symbol of the country’s housing crisis when she was evicted two weeks ago, prompting mass protests and triggering national elections.

中文摘要 María del Carmen Abascal,因被驱逐而成为西班牙住房危机的象征,87岁去世。

What is the pneumonic plague?

The ancient disease of plague is back in the headlines after the death this month of a Russian lab worker.

中文摘要 俄罗斯西伯利亚实验室一名工人死于肺鼠疫,古老疾病再次成为头条。

‘Neither war nor peace’: Palestinians in Gaza on existing amid disease, despair and ruins

People tell of lives changed beyond recognition after the 7 October attacks by Hamas three years ago and the subsequent Israeli offensive Ahmed Ishtaiwi has six chairs. They are cheap and plastic and battered, but they are among the 43-year-old driver and mechanic’s most valuable possessions. Contin

中文摘要 在三年前的哈马斯袭击和随后以色列的进攻之后,加沙地带的巴勒斯坦人生活在疾病、绝望和废墟之中。

María del Carmen Abascal, symbol of Spain’s housing protests, dies aged 87

Maricarmen’s eviction from Madrid apartment after over seven decades set off wave of protests against housing crisis ‘Now is the moment’: housing activists at Madrid encampment pin wary hopes on Spain’s early election María del Carmen Abascal, the 87-year-old whose eviction set off a nationwide wave

中文摘要 Maricarmen Abascal因被驱逐出居住70多年的公寓而引发西班牙住房抗议,87岁去世。

WHO says it does not have ‘full picture’ from Russia following suspected plague death

‘Transparent information sharing’ from Moscow is required to conduct full risk assessment, UN agency says The World Health Organization has said it does not have the “full picture” from Russia following the suspected death from plague last week of a 28-year-old laboratory worker in Siberia. Darya Sh

中文摘要 世界卫生组织表示,由于俄罗斯未能提供完整信息,无法对疑似鼠疫死亡事件进行全面风险评估。

A Plague Mystery in Siberia

The death of a Russian lab worker has spawned rumors. Here are the facts.

中文摘要 俄罗斯西伯利亚实验室工人死亡引发关于肺鼠疫的谣言。

Latest Oil Market News and Analysis for DATE

Oil edged higher after a report that the White House asked the Pentagon to draw up strike options against Iran that could be executed before the midterm elections, and as a storm shut some US output.

中文摘要 受白宫要求五角大楼制定针对伊朗的打击方案以及美国部分油田因风暴关闭的影响,油价上涨。

Braskem Creditors Balk at Restructuring Plan as Deadline Looms

Braskem’s creditors shunned another restructuring proposal made by the company and its shareholders ahead of a key Oct. 9 deadline for a deal to rework some $11 billion of debt, people familiar with the matter said.

中文摘要 Braskem的债权人拒绝了公司及其股东提出的另一项重组计划,该计划旨在重新安排约110亿美元的债务,截止日期为10月9日。

Saudi-Led Forces Claim Ground Advances Against Houthis in Yemen

Yemen’s internationally-recognized government, backed by Saudi Arabia, claimed to have regained territory from the Houthis, as fighting in the region near the Bab el-Mandeb strait intensifies. The Yemeni armed forces said they took control of the Red Sea city of Mokha on Monday after battles with th

中文摘要 也门政府在国际社会支持下,声称从胡塞武装手中夺回领土,冲突加剧。

AI upends Singapore’s ‘quant Olympics’

Nigerian student Victor Ayebameru takes first place in competition that identifies future hedge fund stars

中文摘要 尼日利亚学生Victor Ayebameru在识别未来对冲基金经理竞赛中获得第一名。

Micron Rallies as AI Tailwinds Drive a Street-High Price Target | Closing Bell

Comprehensive cross-platform coverage of the U.S. market close on Bloomberg Television, Bloomberg Radio, and YouTube with Scarlet Fu, Jess Menton, Carol Massar and Lisa Mateo. (Source: Bloomberg)

中文摘要 美光科技在人工智能的推动下股价上涨,收盘价达到历史最高水平。

First Brands’ Brothers Say They’ll Blame Each Other at Trial

The two brothers who built First Brands Group into a global business, and whom prosecutors accuse of driving it into bankruptcy through a multibillion-dollar fraud, are now pointing the finger at each other.

中文摘要 创建First Brands Group的两位兄弟在审判中将相互指责,指责对方将公司带入数十亿美元的欺诈并导致破产。

Fed Minutes Show Hawkish Unity Behind September Rate Hike

Bloomberg's Mike McKee said that the Fed's decision to raise rates was unanimous among all nineteen members of the FOMC, including non-voting participants, according to minutes from the Fed's September meeting. McKee said the minutes showed that members feel that between inflation numbers, the geopo

中文摘要 美联储会议纪要显示,在9月加息决策中,包括非投票参与者在内的所有19名FOMC成员意见一致,显示出鹰派立场。

Midterm Elections Are Potential Test for US Defense Investors

The US midterm elections are a test of investor sentiment in the rapidly growing defense sector, market watchers said.

中文摘要 美国中期选举是对快速增长国防行业投资者情绪的潜在考验。

Why China Shock 2.0 Is Different

And why this period might be more painful for Europe.

中文摘要 探讨中国冲击2.0的不同之处以及这一时期可能对欧洲造成的更大痛苦。

该源今日无内容。