每日简报

2026-09-05

← 历史归档

mattpocock/skills

Shell · ★ 250,284 · 🍴 21,151 · 📈 2,757 stars today

Skills for Real Engineers. Straight from my .agents directory.

中文介绍 为真实工程师提供技能,从作者的 .agents 目录中提取。

DietrichGebert/ponytail

JavaScript · ★ 125,874 · 🍴 6,761 · 📈 1,683 stars today

Makes your AI agent think like the laziest senior dev in the room. The best code is the code you never wrote.

中文介绍 使AI代理像最懒惰的高级开发人员一样思考,最好的代码是你从未写过的代码。

fmtlib/fmt

C++ · ★ 25,456 · 🍴 3,035 · 📈 681 stars today

A modern formatting library

中文介绍 一个现代格式化库,用于代码格式化。

affaan-m/ECC

JavaScript · ★ 248,466 · 🍴 37,444 · 📈 1,139 stars today

The agent harness performance optimization system. Skills, instincts, memory, security, and research-first development for Claude Code, Codex, Opencode, Cursor and beyond.

中文介绍 性能优化系统,为Claude Code、Codex、Opencode、Cursor等提供技能、本能、记忆、安全和研究优先的开发。

anthropics/skills

Python · ★ 174,112 · 🍴 20,639 · 📈 512 stars today

Public repository for Agent Skills

中文介绍 用于存储代理技能的公共仓库。

blader/humanizer

Python · ★ 42,665 · 🍴 3,606 · 📈 1,132 stars today

Agent skill that removes signs of AI-generated writing from text

中文介绍 去除文本中AI生成写作迹象的代理技能。

NousResearch/hermes-agent

Python · ★ 241,462 · 🍴 49,548 · 📈 721 stars today

The agent that grows with you

中文介绍 随着你成长而发展的代理。

JuliusBrussee/caveman

Go · ★ 103,559 · 🍴 6,007 · 📈 503 stars today

🪨 why use many token when few token do trick — Claude Code skill that cuts 65% of tokens by talking like caveman

中文介绍 Claude Code技能,通过像穴居人一样说话,减少65%的token使用。

magnitudedev/magnitude

TypeScript · ★ 2,440 · 🍴 172 · 📈 395 stars today

Open source inference server that runs the best local models for your hardware, plugged into the agent you already use. Works with Pi, OpenCode, Hermes, OpenClaw, Codex, Claude Code, Oh My Pi, and Cline.

中文介绍 开源推理服务器,运行最佳本地模型,与现有代理集成。支持Pi、OpenCode、Hermes、OpenClaw、Codex、Claude Code、Oh My Pi和Cli。

bikini/exploitarium

Python · ★ 4,499 · 🍴 1,240 · 📈 68 stars today

A single archive of public exploit PoCs and vulnerability research writeups. At the time I post these, none have been reported. Feel free to report them yourself and take credit for the CVE if handed out lulz. Please do not abuse these. I do this so to allure people into the field, and I've always f

中文介绍 公共漏洞利用PoC和漏洞研究写作的归档。

bannedbook/fanqiang

Kotlin · ★ 52,754 · 🍴 8,522 · 📈 735 stars today

翻墙-科学上网

中文介绍 翻墙工具,实现科学上网。

debpalash/VoiceStudio

Python · ★ 17,894 · 🍴 2,350 · 📈 1,345 stars today

VoiceStudio is the open-source, fully-local ElevenLabs alternative — voice cloning, voice design, video dubbing, dictation, transcription & audiobook creation in 646 languages.

中文介绍 开源的VoiceStudio,作为ElevenLabs的替代品,提供语音克隆、语音设计、视频配音、语音听写、转录和有声书创作等功能。

google-research/timesfm

Python · ★ 31,034 · 🍴 2,956 · 📈 340 stars today

TimesFM (Time Series Foundation Model) is a pretrained time-series foundation model developed by Google Research for time-series forecasting.

中文介绍 Google Research开发的预训练时间序列基础模型,用于时间序列预测。

radixark/miles

Python · ★ 2,545 · 🍴 442 · 📈 55 stars today

Miles is an enterprise-facing reinforcement learning framework for LLM and VLM post-training, forked from and co-evolving with slime.

中文介绍 面向企业的强化学习框架,用于LLM和VLM的后期训练,从slime分支而来,并与slime共同进化。

anomalyco/opencode

TypeScript · ★ 204,095 · 🍴 26,631 · 📈 314 stars today

The open source coding agent.

中文介绍 开源的编码代理。

clshortfuse/renodx

HLSL · ★ 3,517 · 🍴 137 · 📈 759 stars today

Renovation Engine for DirectX Games

中文介绍 DirectX游戏的翻新引擎。

cathrynlavery/diagram-design

HTML · ★ 30,888 · 🍴 1,982 · 📈 426 stars today

38 editorial diagram types for Claude Code, Codex, and Pi. Self-contained HTML + SVG. No shadows. No Mermaid slop.

中文介绍 为Claude Code、Codex和Pi提供38种编辑图表类型,包含自包含的HTML + SVG,无阴影,无Mermaid slop。

PACE: Towards Surfacing Hidden Conflicts in User Requests

👍 16

Personalized assistants should not only comply with user requests but also assess whether those requests are appropriate given the user's current circumstances. However, prior work has primarily focused on accurately executing requests, overlooking the need for assistants to account for context and

WorldReward: Reward Modeling for Camera-Conditioned World Models

👍 15

Camera-conditioned world models generate interactive videos in which commanded actions should induce the expected scene changes while appearance, geometry, and temporal dynamics remain coherent. Existing rewards assess these requirements separately: geometry-based rewards estimate trajectory executi

Random Attention: Rethinking KV Cache Eviction for Efficient Reasoning

👍 123

Large language models achieve superior performance on tasks that require extended reasoning, but long chains of thought make the KV cache a severe memory bottleneck. Existing KV cache compression methods share one paradigm: score each cached token by some estimate of how much it will matter later, a

Editable Visual Design

👍 32

While diffusion base models such as GPT-Image-2 and Nano-Banana exhibit remarkable visual expressiveness, their end-to-end generation inherently yields flattened bitmaps with error-prone text, precluding layer-wise post-editing. Conversely, code-based visual generation via Coding Agents provides pre

Environment Evolution for Terminal Agents

👍 10

Scaling interactive and verifiable environments is critical for training terminal agents. As frontier models become more capable, environments synthesized from scratch become less challenging and thus provide limited learning signals. Recent co-evolution methods iteratively synthesize environments n

Principia: Relational Physics Tests for Video Models

👍 10

Evaluating physical reasoning in video models is difficult because absolute motion measurements depend on frame rate, object scale, and camera calibration, all of which are often ambiguous or unavailable in generated video. We propose a different approach. When two objects in the same scene obey the

Terminal-Universe: Turning Agent Trajectories into Scalable Terminal Environments

👍 213

As terminal-based code agents become prevalent, agent trajectories have accumulated at scale, while realistic, executable environments remain scarce. However, environments are what agent post-training actually requires: each can be re-queried into many verifiable tasks and provides execution feedbac

VeriPhy: Agentic Physical Reasoning for World Model Evaluation and Refinement

👍 2

Visual fluency in generated video does not imply physical reliability, and a scalar quality score alone is incapable of indicating the obligation a clip violates or the moment it fails. We present VeriPhy, an auditable physical-verification system in which a text-only planner compiles the prompt int

Debias-SparseGPT: Bias-Aware Pruning for Large Language Models

👍 2

Model compression techniques such as pruning and quantization facilitate the efficient deployment and acceleration of Large Language Models (LLMs). However, recent studies show that weight sparsification methods, such as SparseGPT, can amplify existing biases in models, with outputs varying signific

Exploring Collaboration between a language and a non-language agent

👍 5

LLMs are increasingly deployed as orchestrators that coordinate specialized subagents to solve complex tasks through natural language. However, in many important domains like game playing and robotics, the strongest available agents are not language models. Integrating non-language agents with LLMs

Replacing Training with Memory: Listwise Selection for Text-to-SQL

👍 2

Modern Text-to-SQL systems often follow generate-execute-select pipelines, generating multiple candidate queries then selecting the best one. Listwise selection, by jointly comparing multiple candidates, has been widely adopted, but fine-tuning listwise selectors is costly. We thus propose a fine-tu

Sparse Readout Prism: Explaining Logit-Lens Scores in Features Instead of Tokens

👍 2

A language model's prediction of its next token develops across layers, and lens methods track this process by decoding intermediate hidden states into tokens. But a lens reading reflects both the hidden state and the readout (the unembedding matrix) used to decode it. Many lenses are fit on a corpu

Using Grounded Theory for Agent Behavior Analysis at Scale

👍 12

Understanding agent behavior requires methods that scale to thousands of trajectories and surface new patterns in long, often unfamiliar tasks where pre-built classifiers fall short. We propose to bring grounded theory into agent trajectory analysis: a six-decade-old qualitative method from the soci

WHALE: A Simple Recipe for Joint Harness-Weight Optimization

👍 29

Agent performance depends jointly on the model parameters and the executable harness code that manages context and control flow. Optimizing either component in isolation can leave the system bottlenecked by its frozen counterpart: weight updates can change which harness is effective, while harness u

FoldingAgent: Inferring Parametric Origami Procedures from Demonstration Videos

👍 3

We present FoldingAgent, an agentic framework for inferring explicit parametric folding programs directly from origami demonstration videos. Our framework leverages the reasoning power of a pre-trained Vision-Language Model (VLM) equipped with a suite of specialized tools that enable the agent to si

Small Language Models as Judges for Rubric-Based Reinforcement Learning

👍 1

Rubric-based reinforcement learning extends RL beyond tasks with exact answers or rule-based verifiers by scoring responses against instance-specific criteria. However, this makes reward computation expensive: training requires repeated rubric judging, often with proprietary APIs or local generative

Locked at the Entrance, Open Inside: Where RLVR Narrows the Solution Space

👍 2

Reinforcement learning with verifiable rewards (RLVR) substantially improves single-sample accuracy (pass@1) but causes the policy's solution space to contract, diminishing the returns of test-time scaling. In this work, we investigate where inside a reasoning trajectory this breadth is lost: does t

The Cognitive Revolution (When Machines Do The Thinking)

@Konstantine · 17.6K 粉丝 · 256.2K 阅 · 529 赞 · 100 转

We externalized our muscles and built the modern world. Now we are externalizing our minds. Most mornings I ride to work in a Waymo. The car is doing physical work: five thousand pounds of metal and

中文介绍 Konstantine探讨认知革命,从肌肉到思维的机械化,并以Waymo为例阐述AI在现实生活中的应用。

CarryGo, here we go!

@CarryGo_AI · 9 粉丝 · 192.8K 阅 · 1.0K 赞 · 2 转

Inside CarryGo lies a world model—helping organizations see, simulate, act and learn as a system. At the product level, CarryGo is a world model, built first for business. At the system level, it is

中文介绍 CarryGo AI发布,介绍其世界模型如何帮助企业和系统层面上的决策与学习。

Prediction: AI will collapse

@wordgrammer · 26.1K 粉丝 · 120.8K 阅 · 602 赞 · 33 转

The dot com bubble got some things right. It noticed that the internet was going to be big. It noticed that the value generated by the internet would roughly scale in proportion with the amount of

中文介绍 wordgrammer预测AI可能面临类似互联网泡沫的危机,强调互联网价值与规模的关系。

How to design a Brand System from 1 reference image

@AmirMushich · 80.6K 粉丝 · 102.2K 阅 · 501 赞 · 42 转

For 10+ years, I've been creating branding & ads for Warner Music, PepsiCo, and others. Now I combine that experience with AI to build hi-end visual systems faster. This article shows my complete

中文介绍 AmirMushich分享从单一图像设计品牌系统的经验,结合AI技术加速高端视觉系统构建。

7 Grok Bots for Marketing

@irabukht · 18.4K 粉丝 · 63.3K 阅 · 536 赞 · 28 转

18 things to know and 7 paste-in prompts for running ads, SEO and GEO on Grok Bot. Collected from 1,000+ marketers in our community. Most of us use Grok Bot to run marketing work end to end: Weekly

中文介绍 irabukht推荐7个Grok Bots营销机器人,提供广告、SEO和GEO的提示词和营销策略。

AI Engineering Skills Map: Using coding agents

@AndrewYNg · 1.9M 粉丝 · 48.9K 阅 · 918 赞 · 136 转

A key AI engineering skill is using coding agents. Your skill at steering them both to write code and to carry out non-code tasks, such as analyzing data or managing system operations, allows you to

中文介绍 AndrewYNg讨论AI工程技能,重点介绍使用编码代理进行代码编写和非代码任务的技巧。

Grok Bot for Work

@joshkim · 8.3K 粉丝 · 42.4K 阅 · 525 赞 · 41 转

I can't do my job without Grok Bot. It happened really quickly. A few weeks ago, I wasn’t leveraging agents like the AI-pilled experts you see on X and LinkedIn. I was still a human in the loop:

中文介绍 joshkim分享Grok Bot在日常工作中的重要性,从依赖人工转向依赖AI代理。

Codex has changed my life as a solopreneur

@jonnym1ller · 20.7K 粉丝 · 36.0K 阅 · 510 赞 · 23 转

It's difficult to exaggerate how absurdly life-changing Codex has been for me as a solopreneur in recent months. And yet people have been saying 'but what do you actually use it for' (so I decided to

中文介绍 jonnym1ller评价Codex对他作为自由职业者的生活影响,并分享实际应用场景。

Introducing Gemini 3.8 Flash and 3.8 Flash Cyber

@GoogleAIStudio · 198.5K 粉丝 · 26.2K 阅 · 507 赞 · 59 转

Our newest Gemini models deliver next-generation intelligence for agentic workflows and cybersecurity. Building on the momentum of 3.7 Flash from three weeks ago and marking our third Flash release in

中文介绍 GoogleAIStudio介绍Gemini 3.8 Flash和3.8 Flash Cyber模型,为工作流和网络安全提供下一代智能。

From rg to zg: Local Search Beyond Keywords

@QwenDevs · 10.9K 粉丝 · 21.4K 阅 · 509 赞 · 55 转

Summary: The information humans and agents need is often scattered across large numbers of local files, making it difficult to locate accurately and efficiently. zg (zvec-grep) is local-first search

中文介绍 QwenDevs介绍zg(zvec-grep)本地搜索工具,帮助快速准确地定位大量本地文件中的信息。

Architecting memory and storage in the AI era

The era of AI inference has arrived. Imagine a healthcare system analyzing millions of data points in real time to accelerate life-saving medical research, or an intelligent assistant instantly resolving thousands of complex customer needs at once. These real-world breakthroughs rely on advanced inf

中文介绍 AI推理时代来临,医疗系统可实时分析百万数据点,智能助手可即时解决数千个复杂客户需求。

Data from drones in Ukraine is fueling a new Wild West marketplace

Battlefields in Ukraine are littered with the remnants of drones, which are now firmly established as a critical weapon of modern warfare. But behind all that wreckage, there’s a new gold mine for the defense sector. The data drones generate will far outlast the wars in which they are used to fight,

中文介绍 乌克兰战场上的无人机残骸成为新的金矿,为国防部门提供大量数据。

[AINews] GPT-6 Astra: OpenAI’s biggest LLM launch of all time

new SOTA computer use and coding, 2.5x pricier per token, but WAY cheaper per task, less monitorable. overall, a very successful launch of OpenAI’s new frontier model class.

中文介绍 OpenAI发布GPT-6 Astra,价格每令牌高出2.5倍,但任务成本更低。

Daybreak for Frontline Defenders: $1B to protect essential services

OpenAI introduces Daybreak for Frontline Defenders. A $1 billion commitment expands access to frontier cyber AI, training, and support for essential services.

中文介绍 OpenAI投资10亿美元支持前线卫士,扩大关键服务的网络AI、培训和支援。

Playco cut manual fixes 50% prototyping games with GPT-6 Astra

Using GPT-6 Astra, Playco built three themed game prototypes from one grey box foundation and reported 50% fewer manual fixes than with the previous model.

中文介绍 Playco使用GPT-6 Astra将游戏原型制作时间缩短50%。

Legora reviewed 41 documents in minutes with GPT-6 Astra

Legora used GPT-6 Astra to review 41 documents in minutes, find all four planted errors, and improve performance by nearly 40% in this financial-review workflow.

中文介绍 Legora使用GPT-6 Astra在数分钟内审查41份文件,发现所有四处植入的错误,并将性能提高近40%。

Safety overview: GPT-6 Astra

GPT-6 Astra is our most capable broadly deployed model and our first to reach the Critical level of cybersecurity capability under our Preparedness Framework.

中文介绍 GPT-6 Astra是我们最广泛部署的模型,首次达到关键级别的网络安全能力。

9 Days After Flood, 2 Men Are Pulled Alive From Mud-Filled Tunnels in Nepal

One of two workers rescued from a hydroelectric plant said he had recited prayers over and over while trapped deep underground. “How many days has it been?” he asked.

中文摘要 尼泊尔洪水后,两名工人被从泥泞隧道中救出,其中一人被困地下九天。

Mock Republican convention website redirects users to Epstein Files

Democratic lawmakers poke fun at prank website, while Republicans slam 'fake news' and redirect users to real one.

中文摘要 模拟共和党大会网站将用户重定向至爱泼斯坦文件,民主党议员对此进行嘲讽,共和党则指责为‘虚假新闻’。

Mistrial declared in Lindsay Clancy murder case, after jury deadlocks

The mistrial now puts the murder case - and Clancy's future - in limbo as to whether she will be held criminally liable in the deaths of her three kids.

中文摘要 林赛·克拉克谋杀案因陪审团意见分歧而宣布审判无效,案件和克拉克的未来陷入悬念。

Inside the Nepal Tunnel Rescue Operation

An Australian engineer advising the effort described how two men were freed after nine days trapped in a tunnel at the Upper Trishuli 3A hydropower plant.

中文摘要 澳大利亚工程师描述了尼泊尔隧道救援行动,两名男子在被困隧道九天后获救。

U.N. Approves African Proposal for a New World Map

The Equal Earth projection shows countries’ true size relative to one another, unlike the centuries-old Mercator map, which critics say makes Africa look smaller than it is.

中文摘要 联合国批准非洲提出的显示各国真实相对大小的世界地图提案。

Nigel Farage heckled as he denies Reform UK took illegal funding

Reform UK leader Nigel Farage was heckled by protesters on Friday, as he addressed the party’s conference in Birmingham.

中文摘要 改革英国党领袖奈杰尔·法拉奇在伯明翰的党派大会上被抗议者嘘声打断,否认党派接受非法资金。

Will Europe pay the US to provide military aid to Ukraine?

US President Donald Trump laments past financial support for Ukraine's war effort and has vowed to claw the money back.

中文摘要 美国总统特朗普对过去对乌克兰战争努力的财政支持表示遗憾,并誓言收回这笔钱。

Bloom Energy, Illumina, Everpure to Join S&P 500 This Month

Bloom Energy Corp., Illumina Inc. and Everpure Inc. will join the S&P 500 in the latest quarterly rebalance, S&P Dow Jones Indices said Friday.

中文摘要 Bloom Energy、Illumina 和 Everpure 将于本月加入标普500指数。

Acuerdo de Trump con Venezuela inquieta a petroleras de EE.UU.

El presidente Donald Trump se ha jactado de que su amplio acuerdo petrolero con Venezuela, que en la práctica da a Estados Unidos control sobre gran parte de la riqueza petrolera del país, es posiblemente “el mejor acuerdo jamás firmado”.

中文摘要 特朗普与委内瑞拉的石油协议令美国石油企业担忧。

S&P Downgrades Senegal After Government Unveils Debt Rework

S&P Global Ratings cut Senegal’s credit score further into junk territory, saying a distressed debt exchange or default on the country’s foreign-currency commercial debt is “extremely likely” after the government launched a plan to restructure its obligations.

中文摘要 标普下调塞内加尔信用评级,预计该国可能面临债务违约。

Hedge Funds Hike Bullish Oil Bets to May High as Iran War Flares

Hedge funds turned the most bullish on Brent crude since May as a fresh spate of fighting between the US and Iran heightened concerns about prolonged disruptions to energy flows through the Strait of Hormuz.

中文摘要 由于美伊冲突加剧,对冲基金提高对布伦特原油的看涨押注至5月以来的最高水平。

Lutnick to Review Pentagon Deals Involving Cerberus Companies

Commerce Secretary Howard Lutnick has been chosen by the Department of Defense to review any transactions between the Pentagon and companies associated with Cerberus Capital Management, according to a US official.

中文摘要 美国国防部长霍华德·卢特尼克将审查与凯伯斯资本管理公司相关的五角大楼交易。

Wall Street Risk Complex Defies Rate Threat After Jobs Data

The global bond selloff is proving no wrecking ball for Wall Street, repricing the cost of money without setting off the usual scramble out of risk assets.

中文摘要 华尔街风险复杂度在就业数据强劲后抵御利率威胁。

Algorithm Research CEO on Interest Rates Expectations

Ketaki Sharma, founder and CEO of Algorithm Research, says the inflation is the main factor to watch in assessing the the bond market and the Federal Reserve’s next move. She speaks on "Horizons: Middle East and Africa." (Source: Bloomberg)

中文摘要 Algorithm Research首席执行官Ketaki Sharma表示,通胀是评估债券市场和美联储下一步行动的主要因素。

Honey Deuces, Tickets, Merch: US Open Spending

Bloomberg's Lisa Mateo joins Scarlet Fu and Tom Keene on "Bloomberg Money." They discuss how people are spending their money at the US Open. (Source: Bloomberg)

中文摘要 Bloomberg报道了美网期间的消费情况。

Parents Push Teens to Start Investing Earlier Than They Did

Bloomberg's Josyana Joshua joins Scarlet Fu and Tom Keene on "Bloomberg Money." Many parents are opening accounts that let teens learn the basics of investing with their supervision, in hopes of teaching them about investing early. As a result, financial firms have noticed and are rolling out produc

中文摘要 许多父母在监管下为青少年开设账户,让他们学习投资基础,以尽早教授他们投资知识。

Heinekin Highlights Nonalcoholic Beer, Serena Williams in US Open Partnership

Maggie Timoney, CEO of Heinekin USA, said Heinekin's 0.0 nonalcoholic beer has seen success because consumers are looking to moderate their alcohol consumption, even if they're not looking to cut it out entirely. Timoney said that the beverage company wanted to honor the history of the US Open and t

中文摘要 Heineken在美网期间推广无酒精啤酒,并强调其与塞蕾娜·威廉姆斯的合作关系。

US Open Players Make Fashion Statements, Drive Retail

Dana Telsey, CEO & chief research officer at Telsey Advisory Group, joins Scarlet Fu and Tom Keene on "Bloomberg Money." They discuss the fashion statements being made by players at the US Open and how that translates to consumers. (Source: Bloomberg)

中文摘要 Telsey Advisory Group首席执行官Dana Telsey讨论了美网球员的时尚宣言及其对消费者的影响。

该源今日无内容。