每日简报

2026-08-14

← 历史归档

cathrynlavery/diagram-design

HTML · ★ 14,376 · 🍴 859 · 📈 4,504 stars today

29 editorial diagram types for Claude Code. Self-contained HTML + SVG. No shadows, no Mermaid-slop.

中文介绍 提供29种自包含的HTML和SVG编辑图表类型,用于Claude Code,无需阴影和Mermaid辅助。

semantica-agi/semantica

Python · ★ 6,610 · 🍴 696 · 📈 727 stars today

Graph-Native Infrastructure for Context and Accountable AI Systems

中文介绍 为上下文和负责任的人工智能系统提供图原生基础设施。

anthropics/skills

Python · ★ 169,001 · 🍴 20,127 · 📈 383 stars today

Public repository for Agent Skills

中文介绍 存储智能体技能的公共仓库。

cactus-compute/needle

Python · ★ 4,936 · 🍴 333 · 📈 768 stars today

14MB foundation model for tiny devices; phones, wearables, smart home, and robots.

中文介绍 14MB的基础模型,适用于小型设备,如手机、可穿戴设备、智能家居和机器人。

altic-dev/FluidVoice

Swift · ★ 9,840 · 🍴 663 · 📈 187 stars today

Fastest and only macOS Dictation app with on-device STT and custom trained AI enhancement model. A local Wispr Flow alternative. ⭐ helps a ton :) Windows & iOS waitlist open. Linux soon.

中文介绍 macOS上最快的语音输入应用,具有设备端STT和定制训练的AI增强模型,替代Wispr Flow。

unslothai/unsloth

Python · ★ 71,034 · 🍴 6,405 · 📈 354 stars today

Local UI to run and train LLMs and diffusion models, including Qwen3.8, Kimi K3, MiniMax-H3, Gemma 4, DeepSeek-V4, FLUX and more.

中文介绍 本地用户界面,用于运行和训练LLM和扩散模型,包括Qwen3.8、Kimi K3等。

macro-inc/macro

Rust · ★ 2,585 · 🍴 276 · 📈 1,180 stars today

Macro is a unified workspace for teams: email, chat, docs, tasks, agents, calls, and CRM — @-linked together with shared AI memory.

中文介绍 统一的工作空间,整合邮件、聊天、文档、任务、代理、通话和CRM等,通过共享AI记忆链接。

megadose/holehe

Python · ★ 12,409 · 🍴 1,667 · 📈 166 stars today

holehe allows you to check if the mail is used on different sites like twitter, instagram and will retrieve information on sites with the forgotten password function.

中文介绍 检查邮件在不同网站上是否被使用,并在具有忘记密码功能的网站上检索信息。

smicallef/spiderfoot

Python · ★ 20,661 · 🍴 3,314 · 📈 278 stars today

SpiderFoot automates OSINT for threat intelligence and mapping your attack surface.

中文介绍 自动化开源OSINT工具,用于威胁情报和攻击面映射。

NVIDIA-NeMo/Switchyard

Rust · ★ 1,198 · 🍴 108 · 📈 408 stars today

Switchyard lets LLM applications route traffic across models and providers while preserving native OpenAI and Anthropic API compatibility - enabling flexible model selection, benchmarking, and cost/performance optimization.

中文介绍 让LLM应用能够跨模型和提供商路由流量,同时保持原生OpenAI和Anthropic API兼容性。

holaboss-ai/holaOS

TypeScript · ★ 6,570 · 🍴 599 · 📈 380 stars today

Open-source All in One AI agent workspace. Run any agent — Claude Code, Codex — across your tools (100+ integrations + MCP), apps, browser, and files, with shared memory. Built-in models or BYOK.

中文介绍 开源的AI代理工作空间,运行Claude Code、Codex等代理,提供共享记忆和100+集成。

kepano/obsidian-skills

★ 45,708 · 🍴 3,295 · 📈 411 stars today

Agent skills for Obsidian. Teach your agent to use Obsidian CLI and open formats including Markdown, Bases, JSON Canvas.

中文介绍 为Obsidian提供代理技能,教授代理使用Obsidian CLI和Markdown等开放格式。

3b1b/manim

Python · ★ 90,842 · 🍴 7,529 · 📈 204 stars today

Animation engine for explanatory math videos

中文介绍 用于制作解释性数学视频的动画引擎。

msitarzewski/agency-agents

Shell · ★ 145,175 · 🍴 23,483 · 📈 762 stars today

A complete AI agency at your fingertips - From frontend wizards to Reddit community ninjas, from whimsy injectors to reality checkers. Each agent is a specialized expert with personality, processes, and proven deliverables.

中文介绍 提供前端巫师到Reddit社区忍者,从异想天开注入者到现实检查员的全套AI代理。

Lightricks/LTX-2

Python · ★ 8,905 · 🍴 1,406 · 📈 201 stars today

Official Python inference and LoRA trainer package for the LTX-2 audio–video generative model.

中文介绍 LTX-2音频-视频生成模型的Python推理和LoRA训练包。

lightningpixel/modly

TypeScript · ★ 5,391 · 🍴 574 · 📈 221 stars today

Desktop app to generate 3D models from images using local AI — runs entirely on your GPU

中文介绍 桌面应用程序,可使用本地AI从图像生成3D模型,完全在您的GPU上运行。

infiniflow/ragflow

Go · ★ 88,002 · 🍴 10,351 · 📈 473 stars today

RAGFlow is a leading open-source Retrieval-Augmented Generation (RAG) engine that fuses cutting-edge RAG with Agent capabilities to create a superior context layer for LLMs

中文介绍 融合了前沿RAG技术和代理能力的开源RAG引擎,为LLM创建卓越的上下文层。

Hand Visibility Detector: Per-Keypoint Visibility Estimation for Hands

👍 1

Hand Pose Estimation (HPE) is a fundamental technology for various applications such as AR/VR and robotics. In these applications, the visibility of each hand joint in the image is crucial for assessing the reliability of estimation results under occlusion. However, most existing HPE methods output

Persistent Recursive Worlds Enable Autonomous Software Evolution

👍 4

Complex software systems develop over timescales that exceed the lifespan of any individual coding agent. Most agentic software systems preserve continuity through persistent sessions, memories, managers or shared context. We introduce EvoX Genesis (hereafter, Genesis), which instead makes the softw

AI4AI at Test-Time: Strong-to-Weak Capability Transfer via Harnesses

👍 97

Recent work on distillation transfers the capabilities of large models to smaller ones often by updating the latter's parameters, through teacher forcing, on-policy distillation, and related training-time methods. In this paper, we ask whether such transfer can instead occur at test time. We study s

MBA: Multimodal Benchmark and Agents for Real-World Business Ideation

👍 5

Agentic systems powered by large language models (LLMs) have opened new opportunities for business ideation. Yet existing approaches remain confined to a text-only paradigm, despite the inherently multimodal nature of real-world contexts. We thus introduce MBA-Bench, the first multimodal benchmark f

Spark-to-Paper: End-to-End Research Paper Generation as a Composable Skill

👍 176

Turning a research idea into a complete paper requires more than text generation: the system must retrieve literature, design and execute experiments, revise claims according to evidence, produce publication-ready figures, and maintain consistency across a long generation process. We present Spark-t

Self-Evolving Embodied Agents via Skill-Harness Evolution

👍 9

Embodied agents are increasingly built as systems around foundation models, where performance depends not only on model weights but also on the skills, context, action interfaces, and execution harness surrounding the model. While supervised fine-tuning and reinforcement learning can adapt agents to

Agent Safety Should Be a Runtime Contract

👍 3

The dominant paradigm treats AI safety as a property to be instilled during model training via RLHF, DPO, or Constitutional AI. We argue this is structurally insufficient for autonomous agents that execute code, mutate files, send messages, and modify databases. Agent safety should be a runtime cont

InSight-doc: Agentic Visual Perception for Long-Document Understanding

👍 11

Long-document understanding often requires reasoning over many visually rich pages, making inference costly and prone to context rot. In this work, we propose InSight-doc, an agentic visual perception framework that treats visual resolution as an adaptive reasoning-time resource. InSight-doc starts

Parameter Exploration for RLVR via Variational Learning

👍 4

Exploration has been a focus of reinforcement learning research for a long time. Recently, there has been growing evidence that it is also an important ingredient in LLM reinforcement learning recipes that can significantly impact downstream performance. Many existing methods control exploration in

Decoding-Level Taboo: A Diagnostic Stress Test for LLM Robustness

👍 13

Large language model evaluations typically focus on performance under nominal conditions, creating an illusion of capability where models comfortably walk a narrow, highly optimized generation corridor. In real-world deployments, however, complex system prompts, safety guardrails, and structural con

360CityArena: A Realistic Virtual Urban Navigation Benchmark for Embodied Agents

👍 10

We present 360CityArena, a benchmark for evaluating the urban exploration capabilities of embodied agents within a photorealistic environment constructed from 360-degree videos. Existing outdoor benchmarks either lack sufficient photorealism or complexity, resulting in a considerable gap from real-w

NeuPAT: Neuron-aware Plasticity Allocation Tuning for Language-Preserving MLLMs

👍 2

Multimodal expansion of large language models (LLMs) enables new perceptual capabilities but often compromises the language intelligence acquired during pretraining. In this work, we investigate this phenomenon from the perspective of internal adaptation dynamics and discover that neurons in pretrai

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models

👍 2

While Vision-Language-Action (VLA) models have advanced embodied AI, their fundamentally reactive paradigm severely limits performance in partially observable and long-horizon tasks. When restricted to a single wrist-mounted camera, they inevitably suffer from perception forgetting as objects exit t

CoinRAG: Contextualized Information Nugget KV Cache Reuse for Long-Context RAG

👍 6

Recent optimization studies on Retrieval-Augmented Generation (RAG) have exploited chunk-level KV cache reuse to avoid processing long retrieved contexts for higher efficiency, while significant information redundancy and noise still remain in the coarse-grained chunks. This paper optimizes the Pare

The Illusion of Visual Tool-Use: A Causal Audit of Thinking with Images

👍 6

The "thinking-with-images" paradigm equips multimodal LLMs with active visual operations such as crop-and-zoom. However, models using these operations often achieve only marginal or negative gains over direct inference at substantially higher token cost. They may also repeatedly crop irrelevant regi

OpenART: Scaling Agent Red Teaming via Open-Ended Environment Evolution

👍 182

AI agents operate in persistent environments where early state changes can influence decisions far into the future. Unlike conventional language-model interactions, agent behavior is mediated through a shared state that is repeatedly modified and reused across long-horizon workflows. Current safety

Poor Man's Agentic Modeling: Simulating Large LLM-Agent Societies on a Laptop

👍 2

Simulating societies of many large language model (LLM) agents is expensive, yet the questions asked of such simulations are usually macroscopic: phase behaviour, stylised facts, and scaling with the number of agents N, not the cognition of any single agent. We turn a statistical-physics observation

To FDE, or not to FDE?

@thejessezhang · 85.7K 粉丝 · 270.4K 阅 · 510 赞 · 30 转

Two-thirds of our deployment work is now done autonomously by our own product (via Duet). This post is about why that is, our philosophy on building a product + deployment motion, and how that's

中文介绍 探讨产品构建与部署哲学,分享自主部署产品Duet的实践和优势。

Agentic Code Quality

@addyosmani · 408.3K 粉丝 · 175.5K 阅 · 509 赞 · 65 转

For much of human history, we've evaluated code quality via code review: someone reads what you wrote and makes sure it's clean, thoughtful, fast, understandable, and tests well. For agents, that

中文介绍 分析代码审查在智能体中的应用,探讨如何评估智能体代码质量。

Agent Plugins are the future of Agent Skills

@GoogleCloudTech · 1.3M 粉丝 · 56.7K 阅 · 505 赞 · 76 转

Agent Plugins is an open, vendor-neutral standard for packaging Agent Skills and the MCP servers they depend on into one portable folder that any compatible client can load. Google is joining the

中文介绍 介绍Agent Plugins标准,实现智能体技能的便携式打包和加载。

/show-me: compact visual representations for coding agents

@dexhorthy · 30.0K 粉丝 · 42.9K 阅 · 635 赞 · 31 转

tl;dr make your agent converse visually instead of in walls of prose. Lighter and faster than HTML, good enough for most dev-work shaped problems. Coding agents are pretty much unreadable The

中文介绍 介绍新的编码智能体展示工具/show-me,提供视觉化的交流方式。

Intro to Grok Bot

@mattyp · 44.5K 粉丝 · 33.5K 阅 · 651 赞 · 48 转

I’m always chasing better tools. Notes apps, workflows, optimizations, and now, personal agents. It’s an easy trap to fall into - always building the “custom” thing and sacrificing the work as a

中文介绍 介绍Grok Bot,分享个人智能体构建的经验和陷阱。

Grok 4.6 – A field guide

@ericzakariasson · 80.4K 粉丝 · 32.2K 阅 · 719 赞 · 52 转

Grok 4.6 is out! I've used it for a few weeks as my daily driver across the normal mix of coding and knowledge work, and built a few projects with it specifically to push on where it holds up. It's

中文介绍 发布Grok 4.6版本,分享作为日常驱动器和项目开发的体验。

Introducing Gemini 3.7 Flash

@GoogleAIStudio · 191.7K 粉丝 · 30.7K 阅 · 566 赞 · 60 转

Today, we’re building on the progress of our widely used Flash series by introducing Gemini 3.7 Flash, our most intelligent workhorse model yet for coding and agents. This release comes just three

中文介绍 介绍Gemini 3.7 Flash,强调其在编码和智能体领域的智能和效率。

Flock is tightening its rules in response to a growing surveillance backlash

The police-tech giant Flock is announcing today that it will change officers’ access to its nationwide network of license plate readers, in an apparent effort to quell a growing backlash and win back contracts lost amid concerns about mass surveillance and police abuse. Several changes aim directly

中文介绍 Flock公司为应对日益增长的监控反感和警察滥用问题,调整全国车牌识别网络的使用规则。

The builder’s guide to GPT‑5.6

Learn how startups use GPT-5.6 to build faster, more cost-efficient AI agents with smarter model selection and new Responses API capabilities.

中文介绍 OpenAI发布GPT-5.6构建指南,介绍如何使用GPT-5.6构建更高效的人工智能代理。

Previewing Ultrafast mode: GPT-5.6 Sol at up to 14X the speed

Preview Ultrafast, a new OpenAI API service tier that runs GPT-5.6 Sol up to 14× faster. Powered by Cerebras, it delivers up to 750 output tokens per second.

中文介绍 OpenAI推出Ultrafast模式,GPT-5.6 Sol运行速度提高至14倍。

OpenAI appoints Dali Rajic as Chief Revenue Officer

OpenAI appoints Dali Rajic as Chief Revenue Officer to lead its global revenue organization and help businesses realize the full value of AI.

中文介绍 OpenAI任命Dali Rajic为首席营收官,领导全球营收组织。

How kids feel about AI, in their own words

When we set out to talk to kids about artificial intelligence, we thought we knew what we’d hear. We expected some to tell us they were using it to cheat a little, the way Millennials and Gen Xers opened up CliffsNotes or programmed formulas into their TI-82s, and others to share inspiring ways they

中文介绍 MIT Tech Review AI发表文章,探讨孩子们对人工智能的看法。

[AINews] SpaceXAI Grok 4.6 and Grok @Bot

AI teammate category just had its most significant new entrant yet

中文介绍 SpaceXAI推出Grok 4.6和Grok @Bot,成为AI团队成员类别的新成员。

Scaling AI agents with trustworthy data

Business and technology leaders need no convincing that the time of agentic AI is here. Organizations are rapidly adopting agents, and few executives doubt the technology’s potential to transform work. But many organizations find that realizing the desired return on investment (ROI) from AI hinges o

中文介绍 MIT Tech Review AI发表文章,探讨如何通过可靠数据扩展AI代理。

Putting sign language AI into users’ hands

Introducing sign-language-to-text (SL2T), our breakthrough model powering new sign language features for Deaf and hard of hearing users.

中文介绍 DeepMind推出手语AI模型SL2T,为听障和听力受损用户提供新的手语功能。

[AINews] How to steal a Reasoning Trace

Speculative Decoding by any other name would distil as sweet

中文介绍 Latent Space发布文章,探讨如何窃取推理轨迹。

After 2 Plasma Donor Deaths, Company Pauses Clinics in Canada

Grifols, a Spanish company, said “unfounded attention” on its blood plasma business in Canada had led it to temporarily end collections at 17 clinics.

中文摘要 西班牙公司Grifols因加拿大血浆业务受关注,暂时关闭17家诊所。

Australia news live: Albanese says he argued for tariff exemption in ‘substantial’ call with Trump

PM says he also spoke with the US president about Aukus, which ‘remains full steam ahead’. Follow today’s news live Get our breaking news email, free app or daily news podcast Albanese says gambling inducements are ‘over the top’ Albanese said gambling inducements are “over the top and need to be wo

中文摘要 澳大利亚总理Albanese表示在与特朗普的通话中支持关税豁免,并讨论了Aukus计划。

Putin Visits Islands Seized From Japan in World War II, Angering Tokyo

Prime Minister Sanae Takaichi of Japan has been a public critic of Mr. Putin, who has denounced her country’s sanctions against Russia over the Ukraine war.

中文摘要 普京访问二战被日本占领的岛屿,引发东京愤怒,日本首相Takaichi公开批评普京。

Train Derails Near Town of Lewes in East Sussex, U.K.

Emergency services said they were working at the scene of a derailment near Lewes, in East Sussex, on Thursday. Eighteen other people were less badly hurt.

中文摘要 英国东苏塞克斯郡莱韦斯镇附近火车脱轨,造成18人受伤。

Female Afghans Erased From Public Life in Five Years of Taliban Rule

Decrees restricting women’s rights to study, work, travel and act independently now threaten to damage Afghanistan permanently, experts say.

中文摘要 阿富汗女性在塔利班统治下五年内从公共生活中消失,专家表示这威胁到阿富汗的永久性损害。

Mexico says Colombia rejected its earthquake rescue team

Colombia denies politics played a role after Mexico said its military rescuers were refused entry.

中文摘要 墨西哥表示哥伦比亚拒绝其地震救援队入境,哥伦比亚否认政治干预。

What’s at stake in Zambia’s elections?

The economy dominates voter concerns in the southern African country.

中文摘要 赞比亚选举中,经济是选民关注的主要议题。

Gold Steadies as Traders Weigh Fed Rate Path and Mideast Tension

Gold was little changed after retreating below $4,400 an ounce, as traders weighed the Federal Reserve’s interest-rate path and prospects for a deal to reopen the Strait of Hormuz.

中文摘要 金价在跌破每盎司4400美元后微幅波动,交易者权衡美联储利率路径和霍尔木兹海峡重新开放的可能性。

Japan’s Inflation Paradox Is Creating Winners and Losers

Rising prices are fueling a backlash against the government even as they signal a new era of economic dynamism after years of deflation. Bloomberg's Shery Ahn has more. (Source: Bloomberg)

中文摘要 日本通胀悖论正在创造赢家和输家,物价上涨加剧了政府反对,但同时也预示着经济新活力时代的到来。

20 people injured in UK train derailment

Three carriages on the Southern Rail service are flipped on to their side

中文摘要 英国南部铁路服务发生脱轨事故,20人受伤。

Asian Stocks Poised to Gain, Crude Holds Decline: Markets Wrap

Asian stocks were set to extend gains Friday as further evidence of moderating US inflation and a pullback in oil prices reinforced bets that the Federal Reserve will refrain from raising interest rates next month.

中文摘要 亚洲股市预计将因美国通胀降温而上涨,原油价格回落,市场预期美联储下月不会加息。

Reddit Shares Surge on S&P 500 Inclusion Later This Month

Reddit Inc. will join the S&P 500 next week as part of an off-cycle change, S&P Dow Jones Indices said Thursday, sending the social networking platform’s shares surging more than 10% after the bell.

中文摘要 Reddit公司将加入标准普尔500指数,下周起股价上涨超过10%。

Latest Oil Market News and Analysis for Aug. 14

Oil held a decline as traders monitored efforts toward a deal that could reopen the Strait of Hormuz, while fresh attacks on tankers and energy infrastructure kept the market on edge.

中文摘要 油价下跌,交易者关注霍尔木兹海峡重新开放的努力,同时油轮和能源基础设施的新袭击使市场保持紧张。

Daron Acemoglu: We are "at the Cusp of Losing" Liberal Democracy

A new book 'What Happened to Liberal Democracy?' by Nobel Laureate Daron Acemoglu examines the evolution of liberalism, and its future amid rapidly advancing technologies like AI. He joins Buisnessweek Daily to discuss why he wrote this book and how social media and online connections influence what

中文摘要 诺贝尔奖得主达龙·阿西莫格鲁在《什么是自由民主的过去?》一书中探讨了自由主义的演变及其在人工智能等快速发展的技术面前的未来。

SEC Delays Crypto Regulation Meeting in Latest Industry Setback

The Securities and Exchange Commission canceled a Friday meeting where the agency was expected to unveil new plans for digital assets as landmark crypto legislation remains stalled in Congress.

中文摘要 美国证券交易委员会推迟了加密货币监管会议,这是该行业的最新挫折。

Judge Orders Kalshi to Stop Offering Most Wagers in Washington

Kalshi was told by a judge to stop offering most of its prediction market contracts in Washington state, after regulators said they likely constituted an illegal gambling operation.

中文摘要 法官下令Kalshi停止在华盛顿州提供大多数预测市场合约,监管机构称这些可能构成非法赌博行为。

FirstFT: The battle for control of India’s biggest conglomerate

Also in today’s newsletter: Anthropic investors bet on $2tn valuation and Sanae Takaichi slams Vladimir Putin’s visit to disputed Pacific islands

中文摘要 今日FirstFT:投资者押注印度最大集团价值2万亿美元,Sanae Takaichi批评普京访问有争议的太平洋岛屿。