每日简报

2026-08-25

← 历史归档

Alishahryar1/free-claude-code

Python · ★ 48,932 · 🍴 7,992 · 📈 889 stars today

Use Claude Code, Codex, Pi, and OpenCode for free (1.3B+ free tokens) from your terminal, app, IDE, or phone like OpenClaw (voice supported + ToS friendly)

中文介绍 提供Claude Code、Codex、Pi和OpenCode免费使用(1.3B+免费令牌),支持终端、应用、IDE和手机,如OpenClaw(语音支持+遵守服务条款)

openai/codex

Rust · ★ 117,015 · 🍴 17,841 · 📈 1,990 stars today

Lightweight coding agent that runs in your terminal

中文介绍 轻量级编码代理,可在终端运行

MadsLorentzen/ai-job-search

Python · ★ 34,030 · 🍴 11,885 · 📈 378 stars today

The job search that runs on your machine. AI job application framework built on Claude Code: evaluate postings, tailor CVs, write cover letters, prep interviews. Fork it and own it.

中文介绍 在您的机器上运行的职位搜索。基于Claude Code的人工智能职位申请框架:评估职位,定制简历,撰写求职信,准备面试

multica-ai/andrej-karpathy-skills

★ 206,482 · 🍴 21,083 · 📈 491 stars today

A single CLAUDE.md file to improve Claude Code behavior, derived from Andrej Karpathy's observations on LLM coding pitfalls.

中文介绍 一个CLAUDE.md文件,用于改进Claude Code的行为,源自Andrej Karpathy对LLM编码陷阱的观察

makeplane/plane

TypeScript · ★ 57,909 · 🍴 5,487 · 📈 268 stars today

🔥🔥🔥 Open-source Jira, Linear, Monday, and ClickUp alternative. Plane is a modern project management platform to manage tasks, sprints, docs, and triage.

中文介绍 开源的Jira、Linear、Monday和ClickUp替代品。Plane是一个现代化的项目管理平台,用于管理任务、冲刺、文档和问题分类

NousResearch/hermes-agent

Python · ★ 235,779 · 🍴 47,572 · 📈 899 stars today

The agent that grows with you

中文介绍 随着您的发展而成长的代理

anthropics/claude-plugins-community

Python · ★ 1,339 · 🍴 151 · 📈 490 stars today

Community plugin marketplace for Claude Cowork and Claude Code. Read-only mirror — submit plugins at clau.de/plugin-directory-submission.

中文介绍 Claude Cowork和Claude Code的社区插件市场。只读镜像 - 在clau.de/plugin-directory-submission提交插件

AprilNEA/OpenLogi

Rust · ★ 15,843 · 🍴 429 · 📈 1,102 stars today

⚡️A native, local-first alternative to Logitech Options+, written in Rust 🦀 — remap buttons, DPI, and SmartShift over HID++. No account, no telemetry.

中文介绍 Logitech Options+的本地优先替代品,使用Rust编写🦀 - 通过HID++重新映射按钮、DPI和SmartShift,无需账户,无遥测

apache/maka

TypeScript · ★ 2,887 · 🍴 305 · 📈 408 stars today

Apache Maka (Incubating) is a local-first AI agent workspace. Model messages, tool calls, tool results, permission decisions, and termination events are recorded as an append-only log.

中文介绍 Apache Maka(孵化中)是一个本地优先的人工智能代理工作空间。模型消息、工具调用、工具结果、权限决策和终止事件被记录为追加日志

PostHog/posthog

Python · ★ 38,987 · 🍴 3,273 · 📈 106 stars today

🦔 PostHog is the leading platform for building self-driving products. Our developer tools – AI observability, analytics, session replay, flags, experiments, error tracking, logs, and more – capture all the context agents need to diagnose problems, uncover opportunities, and ship fixes. Steer it all

中文介绍 PostHog是一个领先的构建自动驾驶产品的平台。我们的开发工具 - AI可观察性、分析、会话回放、标志、实验、错误跟踪、日志等 - 捕获数据

openclaw/openclaw

TypeScript · ★ 387,433 · 🍴 81,347 · 📈 160 stars today

Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞

中文介绍 您自己的个人AI助手。任何操作系统。任何平台。龙虾的方式。🦞

AgriciDaniel/claude-obsidian

Python · ★ 11,870 · 🍴 1,327 · 📈 272 stars today

Self-organizing AI second brain for Obsidian + Claude Code. Drop any source and Claude reads, links, and files it into one connected knowledge graph of plain Markdown you own. AI note-taking, personal knowledge management (PKM), and an open-source Notion alternative. Based on Karpathy's LLM Wiki pat

中文介绍 为Obsidian和Claude Code构建的自组织AI第二大脑。将任何来源的内容拖放进去,Claude会读取、链接并归档到您拥有的一个连接的知识图谱中。AI笔记记录,个人

rohitg00/ai-engineering-from-scratch

Python · ★ 48,267 · 🍴 8,500 · 📈 330 stars today

Learn it. Build it. Ship it for others.

中文介绍 学习它。构建它。为他人发货。

basecamp/omarchy

Shell · ★ 30,083 · 🍴 3,058 · 📈 1,055 stars today

Beautiful, Modern & Opinionated Linux

中文介绍 美观、现代且具有观点的Linux

tashfeenahmed/freellmapi

TypeScript · ★ 19,763 · 🍴 2,876 · 📈 153 stars today

7.4 billion tokens per month. 34 free LLM providers. 635 free model endpoints. All behind one /v1 endpoint, plus any custom OpenAI-compatible endpoint. Smart routing, automatic failover, encrypted keys. Personal experimentation only.

中文介绍 每月7.4亿令牌。34个免费LLM提供商。635个免费模型端点。所有这些都通过一个/v1端点,以及任何自定义兼容OpenAI的端点。智能路由,自动故障转移,加密密钥

dani-garcia/vaultwarden

Rust · ★ 66,125 · 🍴 3,139 · 📈 176 stars today

Unofficial Bitwarden compatible server written in Rust, formerly known as bitwarden_rs

中文介绍 非官方的Bitwarden兼容服务器,用Rust编写,以前称为bitwarden_rs

freestylefly/awesome-gpt-image-2

JavaScript · ★ 15,457 · 🍴 1,638 · 📈 2,442 stars today

Prompt as Code | GPT-Image2 工业级提示词引擎与模板库,530+ 个案例逆向工程,20+ 套工业级模板,并提炼出Skills,持续更新中

中文介绍 提示词即代码 | GPT-Image2工业级提示词引擎与模板库,530+个案例逆向工程,20+套工业级模板,并提炼出Skills,持续更新中

VoltAgent/awesome-agent-skills

★ 31,854 · 🍴 3,395 · 📈 600 stars today

A curated collection of 1000+ agent skills from official dev teams and the community, compatible with Claude Code, Codex, Gemini CLI, Cursor, and more.

中文介绍 1000+个来自官方开发团队和社区的代理技能精选集合,兼容Claude Code、Codex、Gemini CLI、Cursor等

tinyhumansai/openhuman

Rust · ★ 37,248 · 🍴 3,698 · 📈 515 stars today

Your Personal AI super intelligence. A brain that builds a local-first memory of your life, a fantastic orchestrator of agent fleets and workflows, and a deep researcher.

中文介绍 您的个人AI超级智能。一个构建您生活本地优先记忆的大脑,一个令人惊叹的代理舰队和工作流程的协调者,以及一个深度研究者

PhysCaP: Grounding Code-as-Policy Agent with Physics-Informed Exploration

👍 2

We present PhysCaP, a Physics-Informed Code-as-Policy agent for active perception in robotic manipulation. While vision-language-action policies excel at imitating demonstrations, they rely on passive observation and fail to infer latent physical properties critical for manipulation. PhysCaP augment

OmniAssistBench: Assistant-style Interaction Benchmark for Omni-LLMs

👍 26

Recent omni-modal large language models (Omni-LLMs) show great potential as real-time video assistants, which continuously perceive environments and guide users to achieve specific goals. Unlike traditional passive video understanding, interactive assistants should actively combine visual states, us

Towards Faithful Simulation of Human Shopping Behavior

👍 4

Simulating realistic user shopping behavior underpins offline evaluation and reinforcement learning in e-commerce scenarios. While recent LLM- and VLM-based simulators have made encouraging progress, reproducing a real browsing session remains difficult for two reasons. (i) Memory Challenge: a shopp

FlashPrefill V2: Block-Sparse Prefill Attention for Long-Context LLM Serving

👍 17

Long-context modeling is a pivotal capability for Large Language Models, yet the quadratic complexity of attention remains a critical bottleneck, particularly during the compute-intensive prefilling phase. Our previous work, FlashPrefill, mitigates this cost through instantaneous pattern discovery a

Daedalus-150M: A Convolution-Attention Hybrid Designed for CPU Inference

👍 5

Small language models are usually built like large ones and then squeezed onto a CPU afterwards. We did the opposite: we fixed the target first, one user, one token at a time, 4-bit weights, ordinary CPU, and chose the architecture to suit it. The result keeps full attention in only 6 of its 18 bloc

SWE-bench Science: Can Coding Agents Resolve Engineering Tasks in Science?

👍 61

Software increasingly functions as part of the scientific instrument itself, making failures in scientific code capable of compromising not only program behavior but also the evidence underlying scientific conclusions. Yet existing evaluations of coding agents largely emphasize aggregate task succes

Repo0: Design-Driven Zero-to-All Code Generation

👍 18

Large language model agents have made substantial progress in code generation, yet most existing systems assume a predefined repository architecture. This assumption does not hold in zero-to-all code generation, where an agent must construct an entire software project directly from natural-language

MemTrapBench: Benchmarking Cognitive Traps in LLM Memory Use

👍 31

Memory has become a key component of large language models, enabling them to retain information and learn from long-term interactions. However, existing memory benchmarks mainly evaluate whether information is correctly extracted, stored, and retrieved, while largely overlooking how retrieved memori

EnvHarness: Awakening Static Worlds for Agent Learning

👍 258

LLM agents learn by interacting with environments, yet these environments are hand-built and static: blind to an agent's weaknesses, and quickly left behind as it improves. While recent environment generation methods attempt to address this, they require domain-specific pipelines, rely on expensive

FACET: Preserving Source Intent and Executable State in Terminal Task Synthesis

👍 115

Training terminal agents requires scalable executable supervision, yet synthesizing high-quality terminal tasks remains challenging. Each task couples an instruction, an initialized environment, a reference solution, and an executable verifier; if these artifacts are generated from inconsistent assu

Human-Centric Intelligence in the Era of Foundation Models: A Survey

👍 4

Human-centric intelligence is evolving in the foundation-model era, with growing emphasis on scale, transferability, and general-purpose modeling. Yet it has not fully integrated with foundation models to achieve the comparable progress seen in them. More importantly, recent advances across this bro

Chain-of-Experience for Continual LLM Improvement

👍 8

Humans continuously learn from experience, whereas conventional large language model (LLM) evaluations ignore the models' ability to improve through inference-time interaction. In this paper, we study how LLMs learn from iterative experience at test time, a setting we refer to as Chain-of-Experience

ParaTempo: Efficient Parallel Reasoning via Temporal Confidence

👍 26

Parallel reasoning improves the accuracy and robustness of large reasoning models by exploring multiple solution paths, but its computational cost grows with reasoning depth and branch count. Existing methods for managing these parallel paths typically rely on final-answer consensus, local token con

The Embedder's Dilemma: LLMs Are Better, but at What Cost?

👍 13

Should you replace your text-embedding pipeline with a large language model? We answer this with a controlled, cost-aware comparison of ten LLMs across six families and 26 embedding models (118M to 14B parameters) on 37 tasks spanning classification, semantic textual similarity (STS), clustering, pa

QuoteBench: How Matched Scores Can Hide Command-Path Failures

👍 7

LLM coding agents issue Bash commands through interfaces that may serialize, wrap, and reparse model output. Matched execution scores alone cannot distinguish command-generation errors from failures introduced after generation. QuoteBench measures this boundary with exact final-state validation on 5

SkillEvo: Self-Renewing Evolution Gradients from Multi-Turn Interaction Feedback

👍 30

Agent Skills are today either hand-authored or produced in a single LLM generation pass, and consequently possess no closed loop through which they might improve from the interaction failures they actually cause. Recent work does close this loop, but derives its feedback from single-turn question-an

how i built an SEO/AEO blog engine

@harsehaj · 3.3K 粉丝 · 135.1K 阅 · 507 赞 · 35 转

i proposed and executed an engineering project end-to-end this summer as a growth engineering intern at @browserbase: blogEO, an engine that audits, rewrites, and generates SEO/AEO-tailored blogs on

中文介绍 博主分享了在 Browserbase 担任增长工程实习生期间,如何构建一个 SEO/AEO 博客引擎的全过程。

Deep Dive: The Next Trillion-Dollar Futures Market

@chamath · 2.4M 粉丝 · 81.1K 阅 · 595 赞 · 52 转

“I actually believe a new asset class will be buying futures of compute. We just don’t have enough compute power right now.” - Larry Fink, CEO of BlackRock Recent news moved that prediction closer to

中文介绍 博主探讨了下一个万亿美元的未来市场,引用了 BlackRock 首席执行官关于计算能力资产类别的预测。

17 Skills I Would Install on a Fresh Hermes Setup (10x Powerful Hack)

@FareaNFts · 66.2K 粉丝 · 43.7K 阅 · 534 赞 · 39 转

If I wiped Hermes tomorrow and started from zero, I would not open a blank chat and "figure it out." I would install skills first. A raw Hermes install is smart. A skilled Hermes setup is unfair.

中文介绍 博主介绍了在全新 Hermes 设置中安装的 17 个技能,强调熟练的 Hermes 设置具有不公平的优势。

How to encourage smarter AI use in the classroom

This article is from Making AI Work, MIT Technology Review’s limited-run newsletter examining how to apply LLMs across industries. To receive it in your inbox, sign up here. Chatbots took many schools by surprise upon their release a few years ago. Suddenly, students carried an app in their phones t

中文介绍 MIT Tech Review AI发文探讨如何鼓励更智能地使用AI于课堂,Chatbots在几年前突然在学校普及,引发关注。

Advancing price-performance for developers with GPT‑5.6 in Kiro

GPT‑5.6 is now available in Kiro, helping developers plan, build, review, and test software with better price-performance.

中文介绍 OpenAI推出GPT-5.6,在Kiro平台提供更优的价格性能,助力开发者软件规划、构建、审查和测试。

Kids outlearn AI—and we still don’t know why

People have been talking to each other for at least 100,000 years, as best we can tell. And in all that time, there has been only one thing in the world that could learn a human language to perfect fluency: a human child. Now there are two. Four short years after the release of ChatGPT,…

中文介绍 MIT Tech Review AI文章指出,儿童在语言学习上超越AI,尽管目前尚不清楚原因。

The Evolution of the Agent Harness

Models keep absorbing the harness into their weights — soon, it will be a harness for human attention rather than for the model.

中文介绍 Latent Space文章探讨了模型吸收注意力的演变,未来将更多地关注人类注意力而非模型本身。

Simulation: the new Scaling Law — Joon Sung Park, Simile AI

Simile’s CEO about his journey from the viral Generative Agents to creating 8 Billion Digital Twins of every living human... and why it’s gone from fun exploration to very serious business.

中文介绍 Simile AI的CEO分享了从生成代理到创建80亿个数字孪生的旅程,并解释了这一转变从娱乐探索到严肃商业的原因。

not much happened today

**Ox Alpha** emerged as a mystery model with strong coding and agentic performance, likely a **Zhipu/GLM-family** model such as **GLM-5.3 Vision**. Analysts suggest its gains come from post-training and infrastructure improvements rather than sheer size, based on the **743B base** of **GLM-5.2** wit

中文介绍 Smol AI News报道了Ox Alpha作为一个神秘模型的出现,其强大的编码和代理性能可能来自Zhipu/GLM-family模型如GLM-5.3 Vision。

Debates over AI consciousness are a trap

“Runaway” AI, “rogue” agents, and “autonomous” actors—the current rhetoric would have you believe that AI agents are not only awake and aware, but angry at their creators. Prominent tech leaders such as Demis Hassabis, Dario Amodei, and Sam Altman push for regulation of these seemingly “superhuman”

中文介绍 MIT Tech Review AI文章指出,关于AI意识的辩论是一个陷阱,不应过分解读AI代理的觉醒和意识。

Iran faces 'greatest financial offensive ever', says US treasury secretary

Scott Bessent says the US will sever all economic ties with the country and that any nation partnering with Iran financially will also be isolated.

中文摘要 美国财政部长称,美国将切断与伊朗的所有经济联系,并与任何与伊朗金融合作的伙伴隔离。

Someone Was Impersonating Collien Fernandes Online. She Says It Was Her Husband.

The German actress Collien Fernandes set out to discover who was pretending to be her and sending men sexually explicit material. Her answer ignited a feminist movement.

中文摘要 德国女演员科利恩·费尔南德斯发现其丈夫在网络上冒充她发送色情材料,引发了一场女权主义运动。

Australia news live: WA to get Blue Poles in painting’s first interstate trip in two decades; Aria bans AI music from charts

Anthony Albanese will announce a loan of the work to the Art Gallery of Western Australia from April to September 2027. Follow today’s news live Get our breaking news email, free app or daily news podcast Number of NSW households in urgent need of priority housing hits highest level in over a year T

中文摘要 澳大利亚总理阿尔班塞宣布,将于2027年4月至9月将画作借给西澳大利亚艺术馆,新南威尔士州有紧急住房需求的家庭数量达到历史最高水平。

How US sanctions on Iran ripple through global markets and consumers

New sanctions hit Iran's aviation, tech, and shipping sectors, amplifying pressure on global markets and energy prices.

中文摘要 美国对伊朗的新制裁打击了伊朗的航空、技术和航运部门,加剧了对全球市场和能源价格的压力。

US removes Syria from ‘state sponsor of terrorism’ list

Washington also rescinded designation of HTS, formerly led by President Ahmed al-Sharaa, as a 'terrorist' organisation.

中文摘要 美国从“恐怖主义国家赞助者”名单上除名了叙利亚,同时取消了 HTS 组织的恐怖组织称号。

Six months on, U.S. Navy families who evacuated Bahrain remain in limbo

When the U.S.-Iran war began, more than 1,000 dependents of U.S. Navy personnel in Bahrain were ordered to evacuate. After six months in temporary conditions, they want to know what the plan is.

中文摘要 美国海军家属在巴林被疏散六个月后仍处于困境,他们想知道未来的计划。

UK Prime Minister Burnham Arrives in Ukraine as Russia Amps Up Warnings

In his first foreign trip since taking office, the British prime minister pledged to stand by Kyiv despite new warnings of potential consequences.

中文摘要 英国首相伯恩ham访问乌克兰,尽管俄罗斯发出新的警告,他承诺支持基辅。

US strikes a vessel in the eastern Pacific, killing 2 people in first attack in months

US Southern Command has struck another vessel in the eastern Pacific, killing two accused of trafficking drugs The US military has launched a deadly strike on a vessel in the Pacific Ocean, killing two people accused of drug trafficking. On Monday, US Southern Command (SouthCom) announced it “execut

中文摘要 美国在南太平洋打击一艘船只,造成两人死亡,这是数月来的首次攻击。

Iconic Palestinian artist Sliman Mansour passes away at age 79

Palestinian artist Sliman Mansour, whose work became a symbol of struggle and resistance, has died at age 79.

中文摘要 巴勒斯坦艺术家斯利曼·曼索尔去世,享年79岁,他的作品成为斗争和抵抗的象征。

Iran Pledges to Defy Trump’s Economic Sanctions

Analysts say Tehran could intensify the dispute militarily after the United States announced new efforts to squeeze Iran’s economy. One Iranian official vowed that “not a single drop of oil” would leave the gulf.

中文摘要 伊朗誓言要违抗特朗普的经济制裁,分析师称德黑兰可能军事化争议。

Iran faces 'greatest financial offensive ever', says US treasury secretary

Scott Bessent says the US will sever all economic ties with the country and that any nation partnering with Iran financially will also be isolated.

中文摘要 美国财政部长称,美国将对伊朗实施前所未有的金融进攻,切断与该国的所有经济联系,并孤立任何与伊朗金融合作的伙伴。

Trump says US to increase tariffs on Canadian cars to 50%

Move marks further escalation of Washington’s trade war with Ottawa

中文摘要 特朗普宣布将对加拿大汽车提高关税至50%,标志着华盛顿与渥太华贸易战的进一步升级。

Sammons Distances Itself From Guggenheim in Call With Lenders

Sammons Financial Group reiterated to lenders that it has some distance from Mark Walter’s Guggenheim Partners, after a report last week focusing on the longstanding ties between the firms hit the value of the life insurer’s bonds.

中文摘要 Sammons金融集团在与债权人的通话中重申,它与Mark Walter的Guggenheim Partners存在距离,上周关于两家公司长期关系的报道影响了该人寿保险公司债券的价值。

New 10p coin enters circulation - will you spot one?

The coin has been designed to feature King Charles III and a Scottish grouse close to extinction.

中文摘要 新发行的10便士硬币进入流通,其设计展示了查尔斯三世和濒临灭绝的苏格兰松鸡。

Shein’s IPO pitch should be more Meta than H&M

Its rivals are not just clothes retailers but any business that provides screen-addicted users with a dopamine hit

中文摘要 Shein的IPO计划更像Meta而非H&M,其竞争对手不仅是服装零售商,还包括任何为屏幕上瘾用户提供多巴胺冲动的业务。

SEC subpoenas Wall Street banks over Situational Awareness

Leopold Aschenbrenner’s AI-focused hedge fund nearly collapsed in July before striking a deal with Citadel

中文摘要 美国证券交易委员会对华尔街银行发出传票,调查其情境意识问题,AI对冲基金在7月几乎崩溃后与Citadel达成协议。

Asian Stocks Under Pressure After US Tech Selloff: Markets Wrap

Asian stocks were set to fall after Wall Street gauges retreated as losses in technology shares overshadowed a decline in oil prices.

中文摘要 亚洲股市在美股科技股抛售后承压,油价下跌的降幅被科技股的亏损所掩盖。

Latest Oil Market News and Analysis for Aug. 25

Oil held a decline as the US ramped up economic pressure on Iran and its trading partners in a bid to force the resumption of energy flows through the Strait of Hormuz.

中文摘要 美国加大了对伊朗及其贸易伙伴的经济压力,油价下跌,试图迫使通过霍尔木兹海峡的能源流动恢复。

FirstFT: Shein targets $27bn valuation in Hong Kong IPO

Also in this newsletter: US-Canada trade war erupts and Singapore’s effort to boost birth rates

中文摘要 Shein计划在香港IPO中实现270亿美元的估值,同时美国与加拿大贸易战爆发,新加坡努力提高出生率。

NZ Market Watchdog Finds Māori Face Insurance Gaps, Cost Barrier

New Zealand’s insurance industry is failing to consistently meet the needs of Māori, with affordability barriers and difficulties insuring collectively owned land leaving them less likely to have cover, according to a study by the market watchdog.

中文摘要 新西兰市场监管机构研究发现,毛利人面临保险差距和成本障碍,可负担性障碍和集体土地保险困难导致他们获得保险的可能性较低。

该源今日无内容。