每日简报

2026-08-14

← 历史归档

cathrynlavery/diagram-design

HTML · ★ 14,210 · 🍴 854 · 📈 4,504 stars today

29 editorial diagram types for Claude Code. Self-contained HTML + SVG. No shadows, no Mermaid-slop.

semantica-agi/semantica

Python · ★ 6,563 · 🍴 692 · 📈 727 stars today

Graph-Native Infrastructure for Context and Accountable AI Systems

anthropics/skills

Python · ★ 168,975 · 🍴 20,129 · 📈 383 stars today

Public repository for Agent Skills

cactus-compute/needle

Python · ★ 4,917 · 🍴 332 · 📈 768 stars today

14MB foundation model for tiny devices; phones, wearables, smart home, and robots.

altic-dev/FluidVoice

Swift · ★ 9,826 · 🍴 662 · 📈 187 stars today

Fastest and only macOS Dictation app with on-device STT and custom trained AI enhancement model. A local Wispr Flow alternative. ⭐ helps a ton :) Windows & iOS waitlist open. Linux soon.

unslothai/unsloth

Python · ★ 71,016 · 🍴 6,404 · 📈 354 stars today

Local UI to run and train LLMs and diffusion models, including Qwen3.8, Kimi K3, MiniMax-H3, Gemma 4, DeepSeek-V4, FLUX and more.

macro-inc/macro

Rust · ★ 2,567 · 🍴 273 · 📈 1,180 stars today

Macro is a unified workspace for teams: email, chat, docs, tasks, agents, calls, and CRM — @-linked together with shared AI memory.

megadose/holehe

Python · ★ 12,388 · 🍴 1,667 · 📈 166 stars today

holehe allows you to check if the mail is used on different sites like twitter, instagram and will retrieve information on sites with the forgotten password function.

smicallef/spiderfoot

Python · ★ 20,644 · 🍴 3,313 · 📈 278 stars today

SpiderFoot automates OSINT for threat intelligence and mapping your attack surface.

NVIDIA-NeMo/Switchyard

Rust · ★ 1,181 · 🍴 104 · 📈 408 stars today

Switchyard lets LLM applications route traffic across models and providers while preserving native OpenAI and Anthropic API compatibility - enabling flexible model selection, benchmarking, and cost/performance optimization.

holaboss-ai/holaOS

TypeScript · ★ 6,534 · 🍴 598 · 📈 380 stars today

Open-source All in One AI agent workspace. Run any agent — Claude Code, Codex — across your tools (100+ integrations + MCP), apps, browser, and files, with shared memory. Built-in models or BYOK.

kepano/obsidian-skills

★ 45,669 · 🍴 3,295 · 📈 411 stars today

Agent skills for Obsidian. Teach your agent to use Obsidian CLI and open formats including Markdown, Bases, JSON Canvas.

3b1b/manim

Python · ★ 90,832 · 🍴 7,529 · 📈 204 stars today

Animation engine for explanatory math videos

msitarzewski/agency-agents

Shell · ★ 145,159 · 🍴 23,483 · 📈 762 stars today

A complete AI agency at your fingertips - From frontend wizards to Reddit community ninjas, from whimsy injectors to reality checkers. Each agent is a specialized expert with personality, processes, and proven deliverables.

Lightricks/LTX-2

Python · ★ 8,896 · 🍴 1,405 · 📈 201 stars today

Official Python inference and LoRA trainer package for the LTX-2 audio–video generative model.

lightningpixel/modly

TypeScript · ★ 5,362 · 🍴 574 · 📈 221 stars today

Desktop app to generate 3D models from images using local AI — runs entirely on your GPU

infiniflow/ragflow

Go · ★ 87,983 · 🍴 10,348 · 📈 473 stars today

RAGFlow is a leading open-source Retrieval-Augmented Generation (RAG) engine that fuses cutting-edge RAG with Agent capabilities to create a superior context layer for LLMs

Hand Visibility Detector: Per-Keypoint Visibility Estimation for Hands

👍 1

Hand Pose Estimation (HPE) is a fundamental technology for various applications such as AR/VR and robotics. In these applications, the visibility of each hand joint in the image is crucial for assessing the reliability of estimation results under occlusion. However, most existing HPE methods output

Persistent Recursive Worlds Enable Autonomous Software Evolution

👍 4

Complex software systems develop over timescales that exceed the lifespan of any individual coding agent. Most agentic software systems preserve continuity through persistent sessions, memories, managers or shared context. We introduce EvoX Genesis (hereafter, Genesis), which instead makes the softw

AI4AI at Test-Time: Strong-to-Weak Capability Transfer via Harnesses

👍 96

Recent work on distillation transfers the capabilities of large models to smaller ones often by updating the latter's parameters, through teacher forcing, on-policy distillation, and related training-time methods. In this paper, we ask whether such transfer can instead occur at test time. We study s

MBA: Multimodal Benchmark and Agents for Real-World Business Ideation

👍 5

Agentic systems powered by large language models (LLMs) have opened new opportunities for business ideation. Yet existing approaches remain confined to a text-only paradigm, despite the inherently multimodal nature of real-world contexts. We thus introduce MBA-Bench, the first multimodal benchmark f

Spark-to-Paper: End-to-End Research Paper Generation as a Composable Skill

👍 176

Turning a research idea into a complete paper requires more than text generation: the system must retrieve literature, design and execute experiments, revise claims according to evidence, produce publication-ready figures, and maintain consistency across a long generation process. We present Spark-t

Self-Evolving Embodied Agents via Skill-Harness Evolution

👍 9

Embodied agents are increasingly built as systems around foundation models, where performance depends not only on model weights but also on the skills, context, action interfaces, and execution harness surrounding the model. While supervised fine-tuning and reinforcement learning can adapt agents to

Agent Safety Should Be a Runtime Contract

👍 3

The dominant paradigm treats AI safety as a property to be instilled during model training via RLHF, DPO, or Constitutional AI. We argue this is structurally insufficient for autonomous agents that execute code, mutate files, send messages, and modify databases. Agent safety should be a runtime cont

InSight-doc: Agentic Visual Perception for Long-Document Understanding

👍 10

Long-document understanding often requires reasoning over many visually rich pages, making inference costly and prone to context rot. In this work, we propose InSight-doc, an agentic visual perception framework that treats visual resolution as an adaptive reasoning-time resource. InSight-doc starts

ComBodied Agents: a New Paradigm of Human-Centric Agentic AI

👍 181

After an older adult misses a medication dose, a software agent can send another reminder and an embodied agent can bring the medication. Yet neither explains whether the person forgot, is confused, has side effects, or deliberately refused, nor what support is appropriate. This reveals a structural

Parameter Exploration for RLVR via Variational Learning

👍 4

Exploration has been a focus of reinforcement learning research for a long time. Recently, there has been growing evidence that it is also an important ingredient in LLM reinforcement learning recipes that can significantly impact downstream performance. Many existing methods control exploration in

Decoding-Level Taboo: A Diagnostic Stress Test for LLM Robustness

👍 12

Large language model evaluations typically focus on performance under nominal conditions, creating an illusion of capability where models comfortably walk a narrow, highly optimized generation corridor. In real-world deployments, however, complex system prompts, safety guardrails, and structural con

360CityArena: A Realistic Virtual Urban Navigation Benchmark for Embodied Agents

👍 10

We present 360CityArena, a benchmark for evaluating the urban exploration capabilities of embodied agents within a photorealistic environment constructed from 360-degree videos. Existing outdoor benchmarks either lack sufficient photorealism or complexity, resulting in a considerable gap from real-w

NeuPAT: Neuron-aware Plasticity Allocation Tuning for Language-Preserving MLLMs

👍 2

Multimodal expansion of large language models (LLMs) enables new perceptual capabilities but often compromises the language intelligence acquired during pretraining. In this work, we investigate this phenomenon from the perspective of internal adaptation dynamics and discover that neurons in pretrai

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models

👍 2

While Vision-Language-Action (VLA) models have advanced embodied AI, their fundamentally reactive paradigm severely limits performance in partially observable and long-horizon tasks. When restricted to a single wrist-mounted camera, they inevitably suffer from perception forgetting as objects exit t

CoinRAG: Contextualized Information Nugget KV Cache Reuse for Long-Context RAG

👍 5

Recent optimization studies on Retrieval-Augmented Generation (RAG) have exploited chunk-level KV cache reuse to avoid processing long retrieved contexts for higher efficiency, while significant information redundancy and noise still remain in the coarse-grained chunks. This paper optimizes the Pare

The Illusion of Visual Tool-Use: A Causal Audit of Thinking with Images

👍 6

The "thinking-with-images" paradigm equips multimodal LLMs with active visual operations such as crop-and-zoom. However, models using these operations often achieve only marginal or negative gains over direct inference at substantially higher token cost. They may also repeatedly crop irrelevant regi

OpenART: Scaling Agent Red Teaming via Open-Ended Environment Evolution

👍 181

AI agents operate in persistent environments where early state changes can influence decisions far into the future. Unlike conventional language-model interactions, agent behavior is mediated through a shared state that is repeatedly modified and reused across long-horizon workflows. Current safety

Poor Man's Agentic Modeling: Simulating Large LLM-Agent Societies on a Laptop

👍 2

Simulating societies of many large language model (LLM) agents is expensive, yet the questions asked of such simulations are usually macroscopic: phase behaviour, stylised facts, and scaling with the number of agents N, not the cognition of any single agent. We turn a statistical-physics observation

To FDE, or not to FDE?

@thejessezhang · 85.7K 粉丝 · 270.4K 阅 · 510 赞 · 30 转

Two-thirds of our deployment work is now done autonomously by our own product (via Duet). This post is about why that is, our philosophy on building a product + deployment motion, and how that's

中文介绍 探讨产品构建与部署自动化哲学,分享 Duet 工具在自动化部署中的应用。

Agentic Code Quality

@addyosmani · 408.3K 粉丝 · 175.5K 阅 · 509 赞 · 65 转

For much of human history, we've evaluated code quality via code review: someone reads what you wrote and makes sure it's clean, thoughtful, fast, understandable, and tests well. For agents, that

中文介绍 分析代码审查在人类历史中的作用,探讨代码质量在智能体中的应用。

Agent Plugins are the future of Agent Skills

@GoogleCloudTech · 1.3M 粉丝 · 56.7K 阅 · 505 赞 · 76 转

Agent Plugins is an open, vendor-neutral standard for packaging Agent Skills and the MCP servers they depend on into one portable folder that any compatible client can load. Google is joining the

中文介绍 介绍 Agent Plugins 标准,Google 加入推动智能体技能的标准化。

/show-me: compact visual representations for coding agents

@dexhorthy · 30.0K 粉丝 · 42.9K 阅 · 635 赞 · 31 转

tl;dr make your agent converse visually instead of in walls of prose. Lighter and faster than HTML, good enough for most dev-work shaped problems. Coding agents are pretty much unreadable The

中文介绍 提出 /show-me 工具,为编码智能体提供可视化对话界面,提升开发效率。

Intro to Grok Bot

@mattyp · 44.5K 粉丝 · 33.5K 阅 · 651 赞 · 48 转

I’m always chasing better tools. Notes apps, workflows, optimizations, and now, personal agents. It’s an easy trap to fall into - always building the “custom” thing and sacrificing the work as a

中文介绍 介绍 Grok Bot,分享个人智能体构建与工作优化的心得。

Grok 4.6 – A field guide

@ericzakariasson · 80.4K 粉丝 · 32.2K 阅 · 719 赞 · 52 转

Grok 4.6 is out! I've used it for a few weeks as my daily driver across the normal mix of coding and knowledge work, and built a few projects with it specifically to push on where it holds up. It's

中文介绍 发布 Grok 4.6 版本,分享作为日常工具的体验和项目应用。

Introducing Gemini 3.7 Flash

@GoogleAIStudio · 191.7K 粉丝 · 30.7K 阅 · 566 赞 · 60 转

Today, we’re building on the progress of our widely used Flash series by introducing Gemini 3.7 Flash, our most intelligent workhorse model yet for coding and agents. This release comes just three

中文介绍 推出 Gemini 3.7 Flash,强化编码和智能体工作的智能模型。

Introducing Gemini 3.7 Flash

中文介绍 DeepMind Blog推出Gemini 3.7 Flash,提供更快速的数据处理能力。

Flock is tightening its rules in response to a growing surveillance backlash

The police-tech giant Flock is announcing today that it will change officers’ access to its nationwide network of license plate readers, in an apparent effort to quell a growing backlash and win back contracts lost amid concerns about mass surveillance and police abuse. Several changes aim directly

中文介绍 MIT Tech Review AI报道,警察技术公司Flock为平息对大规模监控和警察滥用担忧的反弹,将调整警察对全国车牌识别网络的访问权限。

The builder’s guide to GPT‑5.6

Learn how startups use GPT-5.6 to build faster, more cost-efficient AI agents with smarter model selection and new Responses API capabilities.

中文介绍 OpenAI发布GPT-5.6构建指南,帮助初创公司构建更快速、成本效益更高的AI代理。

Previewing Ultrafast mode: GPT-5.6 Sol at up to 14X the speed

Preview Ultrafast, a new OpenAI API service tier that runs GPT-5.6 Sol up to 14× faster. Powered by Cerebras, it delivers up to 750 output tokens per second.

中文介绍 OpenAI展示Ultrafast模式,GPT-5.6 Sol速度提升至14倍,输出速度达到每秒750个token。

OpenAI appoints Dali Rajic as Chief Revenue Officer

OpenAI appoints Dali Rajic as Chief Revenue Officer to lead its global revenue organization and help businesses realize the full value of AI.

中文介绍 OpenAI任命Dali Rajic为首席营收官,领导全球营收组织。

How kids feel about AI, in their own words

When we set out to talk to kids about artificial intelligence, we thought we knew what we’d hear. We expected some to tell us they were using it to cheat a little, the way Millennials and Gen Xers opened up CliffsNotes or programmed formulas into their TI-82s, and others to share inspiring ways they

中文介绍 MIT Tech Review AI探讨儿童对人工智能的看法,发现孩子们对AI的应用看法多样。

[AINews] SpaceXAI Grok 4.6 and Grok @Bot

AI teammate category just had its most significant new entrant yet

中文介绍 Latent Space报道,SpaceXAI推出Grok 4.6和Grok @Bot,为AI团队增加新的成员。

Scaling AI agents with trustworthy data

Business and technology leaders need no convincing that the time of agentic AI is here. Organizations are rapidly adopting agents, and few executives doubt the technology’s potential to transform work. But many organizations find that realizing the desired return on investment (ROI) from AI hinges o

中文介绍 MIT Tech Review AI讨论使用可信数据进行AI代理扩展的挑战。

Putting sign language AI into users’ hands

Introducing sign-language-to-text (SL2T), our breakthrough model powering new sign language features for Deaf and hard of hearing users.

中文介绍 DeepMind Blog推出手语到文本(SL2T)模型,为听障和听力受损用户提供新的手语功能。

[AINews] How to steal a Reasoning Trace

Speculative Decoding by any other name would distil as sweet

中文介绍 Latent Space分享如何获取推理轨迹的方法。

Australia news live: Albanese says he argued for tariff exemption in ‘substantial’ call with Trump

PM says he also spoke with the US president about Aukus, which ‘remains full steam ahead’. Follow today’s news live Get our breaking news email, free app or daily news podcast Albanese says gambling inducements are ‘over the top’ Albanese said gambling inducements are “over the top and need to be wo

中文摘要 阿尔班斯表示在与特朗普的“实质性”通话中支持关税豁免,并讨论了“Aukus”协议。

Conditions on US aircraft carrier at sea for more than 250 days raise alarms

Thousands of sailors on the USS Abraham Lincoln have reportedly faced food shortages and broken plumbing, with some considering jumping overboard.

中文摘要 美国海军陆战队USS Abraham林肯号航母上的数千名水兵面临食物短缺和破损的管道问题,一些水手考虑跳海。

Train Derails Near Town of Lewes in East Sussex, U.K.

Emergency services said they were working at the scene of a derailment near Lewes, in East Sussex, on Thursday. Eighteen other people were less badly hurt.

中文摘要 周四,在英国东苏塞克斯郡利文斯附近的火车脱轨事件中,18人受轻伤。

Female Afghans Erased From Public Life in Five Years of Taliban Rule

Decrees restricting women’s rights to study, work, travel and act independently now threaten to damage Afghanistan permanently, experts say.

中文摘要 五年来塔利班统治下,阿富汗女性从公共生活中消失,专家警告这可能会永久损害阿富汗。

Mexico says Colombia rejected its earthquake rescue team

Colombia denies politics played a role after Mexico said its military rescuers were refused entry.

中文摘要 墨西哥表示,哥伦比亚拒绝了其地震救援队入境,但哥伦比亚否认政治介入。

What’s at stake in Zambia’s elections?

The economy dominates voter concerns in the southern African country.

中文摘要 赞比亚选举中,经济问题是选民关注的焦点。

Photos: Cuba marks Fidel Castro’s 100th birthday

Fidel Castro would have turned 100 on Thursday and his legacy continues to overshadow the island.

中文摘要 古巴庆祝菲德尔·卡斯特罗100岁生日,其遗产继续笼罩着这个岛国。

Arabian Sea Bird and Turtle Habitat Threatened by Oil From Grounded Tanker

The spill of Russian crude bound for India compounds the environmental damage from spills in and around the Strait of Hormuz, where tankers and oil facilities have come under attack.

中文摘要 俄罗斯原油在驶往印度的途中泄漏,威胁到阿拉伯海鸟类和海龟栖息地,加剧了霍尔木兹海峡及其周边地区的环境损害。

Asian Stocks Set for Gains as US Inflation Cools: Markets Wrap

Asian stocks were poised to extend gains Friday as further evidence of moderating US inflation and a pullback in oil prices reinforced bets that the Federal Reserve will refrain from raising interest rates next month.

中文摘要 亚洲股市预期上涨,因美国通胀降温及油价回落,市场预期美联储下月不会加息。

SEC Delays Crypto Regulation Meeting in Latest Industry Setback

The Securities and Exchange Commission canceled a Friday meeting where the agency was expected to unveil new plans for digital assets as landmark crypto legislation remains stalled in Congress.

中文摘要 美国证券交易委员会推迟加密货币监管会议,标志着行业最新挫折。

Judge Orders Kalshi to Stop Offering Most Wagers in Washington

Kalshi was told by a judge to stop offering most of its prediction market contracts in Washington state, after regulators said they likely constituted an illegal gambling operation.

中文摘要 Kalshi被法官下令停止在华盛顿州提供大部分预测市场合约,监管机构称其可能构成非法赌博。

20 people injured in UK train derailment

Three carriages on the Southern Rail service are flipped on to their side

中文摘要 英国火车脱轨事故造成20人受伤。

Winklevosses’ Gemini Posts Another Loss, Revenue Increases

Gemini Space Station Inc., the digital-asset platform led by the billionaires Tyler and Cameron Winklevoss that went public last year just before the crypto market tumbled from record highs, posted a fourth consecutive quarterly loss since the IPO.

中文摘要 Winklevoss兄弟的Gemini平台发布第四季度连续亏损,但收入增加。

The Anti-Aging Movement Has Come for Cats and Dogs

Supplement stacks. Red light therapy. Peptide shots. Many people are biohacking their health to try to live longer, better lives. And now, they’re doing the same for their beloved pets. Anna Edney reports. (Source: Bloomberg)

中文摘要 抗衰老运动也波及猫狗,许多人尝试通过生物黑客技术延长宠物的寿命。

S&P 500 Hits Record as Inflation Cools | Closing Bell

Comprehensive cross-platform coverage of the U.S. market close on Bloomberg Television, Bloomberg Radio, and YouTube with Romaine Bostick, Isabelle Lee, Carol Massar and Tim Stenovec. (Source: Bloomberg)

中文摘要 标普500指数因通胀降温创下历史新高。

Fmr. NBA Executive: Lakers a ‘Special Asset’ in Sports

Bobby Sharma, Founder and Managing Partner at Bluestone Equity Partners, discussed the recent $12.5 billion valuation of the Los Angeles Lakers, highlighting the extraordinary price and the broader trend of sports becoming institutionalized investment assets. Sharma emphasized that the Lakers repres

中文摘要 前NBA高管:洛杉矶湖人队是体育的特殊资产,其估值达到125亿美元。

UAE, Qatari Firms Debut in Venezuela Through BP-Led Gas Deal

Arab Gulf companies are making their debut in Venezuela through an offshore natural gas project to be operated by BP Plc.

中文摘要 阿联酋和卡塔尔公司通过BP领导的天然气项目在委内瑞拉首次亮相。

Tyson to Close More Beef Plants as US Cattle Shortage Drags On

Tyson Foods Inc. said it would close several additional beef plants, in the latest sign of the US beefpacking industry’s restructuring amid a prolonged US cattle shortage.

中文摘要 泰森食品因美国牛肉短缺持续,将关闭更多牛肉加工厂。

该源今日无内容。