每日简报

2026-10-07

← 历史归档

tester-army/e2e

TypeScript · ★ 6,358 · 🍴 285 · 📈 1,725 stars today

Next generation e2e testing framework for web and mobile apps.

中文介绍 提供下一代端到端测试框架,适用于网页和移动应用,用于提升测试效率和稳定性。

mattpocock/skills

Shell · ★ 278,171 · 🍴 23,291 · 📈 889 stars today

Skills for Real Engineers. Straight from my .agents directory.

中文介绍 为真实工程师提供技能集,源自作者的个人技能目录,可用于提升个人技能管理。

earthtojake/text-to-cad

Python · ★ 17,990 · 🍴 1,805 · 📈 619 stars today

Give your agent CAD superpowers.

中文介绍 赋予代理CAD能力,通过文本转换为CAD设计,适用于需要快速生成CAD模型的场景。

boykopovar/AnyPS5

C++ · ★ 6,591 · 🍴 489 · 📈 949 stars today

Tool for automatic PS5 executables porting to Linux and Windows

中文介绍 自动将PS5可执行文件移植到Linux和Windows系统,适用于需要跨平台运行PS5游戏或应用的开发者。

pbakaus/impeccable

JavaScript · ★ 77,707 · 🍴 4,629 · 📈 616 stars today

The design language that makes your AI harness better at design.

中文介绍 设计语言,通过优化AI设计过程,提升设计效率和效果。

thedotmack/claude-mem

TypeScript · ★ 97,193 · 🍴 8,564 · 📈 534 stars today

Persistent Context Across Sessions for Every Agent – Captures everything your agent does during sessions, compresses it with AI, and injects relevant context back into future sessions. Works with Claude Code, OpenClaw, Codex, Gemini, Hermes, Copilot, OpenCode + More

中文介绍 为每个代理提供会话间的持久上下文,通过AI压缩和注入上下文,适用于需要持续学习上下文的智能代理。

ayghri/i-have-adhd

Python · ★ 54,424 · 🍴 3,125 · 📈 326 stars today

A skill to stop your coding agent from burying the answer. ADHD-friendly output.

中文介绍 专为ADHD患者设计的技能,帮助编码代理避免遗漏答案,提升编码效率。

morluto/rea

TypeScript · ★ 9,526 · 🍴 1,059 · 📈 2,956 stars today

Reverse engineer anything with agents, from app behavior down to native binaries.

中文介绍 使用代理进行逆向工程,从应用行为到原生二进制文件,适用于需要逆向分析各种软件的工程师。

deepseek-ai/DeepGEMM

Cuda · ★ 8,712 · 🍴 1,372 · 📈 199 stars today

DeepGEMM: clean and efficient BLAS kernel library on GPU

中文介绍 在GPU上提供高效且干净的BLAS内核库,适用于需要高性能数学运算的深度学习应用。

msitarzewski/agency-agents

Shell · ★ 157,837 · 🍴 25,448 · 📈 623 stars today

A complete AI agency at your fingertips - From frontend wizards to Reddit community ninjas, from whimsy injectors to reality checkers. Each agent is a specialized expert with personality, processes, and proven deliverables.

中文介绍 提供全面的AI代理服务,包括前端专家、Reddit社区忍者等,适用于需要多样化AI代理的用户。

DuarteSantos8/openGym

JavaScript · ★ 5,719 · 🍴 783 · 📈 1,419 stars today

Self-hosted gym & body-weight tracker — plan routines, log workouts (supersets, warm-ups, cardio), see which muscles are trained, fatigued or detrained, import from FitNotes/Strong/Hevy, passkey login. Your data, your server.

中文介绍 自托管健身和体重跟踪器,支持计划训练、记录锻炼、查看肌肉训练状态,适用于健身爱好者。

cathrynlavery/diagram-design

HTML · ★ 44,048 · 🍴 2,842 · 📈 228 stars today

Editorial diagram design for Claude Code, Codex, GitHub Copilot, Factory Droid, and Pi. 42 diagram types. Self-contained HTML + SVG. No shadows. No Mermaid slop.

中文介绍 为Claude Code、Codex、GitHub Copilot等提供编辑器,支持多种图表类型,适用于需要快速创建图表的设计师。

LoGRA: Scaling LLM Reinforcement Learning with Low-Rank Gradient Sketches

👍 4

Reinforcement learning (RL) has greatly advanced the capabilities of large language models (LLMs), but its memory demands remain a barrier to broader adoption. We introduce LoGRA, an approach to RL post-training that reduces memory by retaining useful learning signals in low-rank gradient sketches.

Capability-Driven Self-Evolution of Agent Memory

👍 2

Memory self-evolution uses task feedback to iteratively improve executable memory programs that store and retrieve information from past interactions. Existing approaches typically adopt holistic evolution, deriving revision directions from mixed feedback and judging progress by overall performance.

Closing the Context Gap: Activation Alignment for Tabular In-Context Learning

👍 1

Tabular foundation models perform in-context learning (ICL) by conditioning predictions on labeled training examples provided as context. Unlike traditional models that separate training from inference, these models must process all training examples in every forward pass, making each prediction exp

Collaborative Personalized Preference Alignment for LLMs under Data Deficiency

👍 2

Real-world users often exhibit highly heterogeneous preferences over multiple objectives for LLM responses. A lightweight aligner can tailor these responses to individual preferences, but scarce user-specific feedback makes personalized training difficult. Learning shared initializations across user

Adapting prior-data fitted networks for tabular anomaly detection

👍 2

While deep features have transformed anomaly detection in images and video, their impact on tabular data has been less substantial, partly due to the limited availability of strong deep representations. Recently, prior-data fitted networks (PFNs) have emerged as a promising source of such representa

SoK: Semantic Decision Engines in Network Control Loops

👍 2

A semantic decision engine such as Jev can return a valid answer and still miss a network deadline, select an infeasible action or leave the service unverified. We systematize 139 paper families by decision interface, execution path and check ownership. Fifty families claim that their engine fits a

What Matters for Latent Reasoning with Flow Matching

👍 6

Latent reasoning lets a large language model (LLM) think in a continuous space and verbalize only the answer. We argue that an effective latent thought must meet five requirements: it should be useful, helping produce the correct answer rather than merely changing it, diverse, so that resampling yie

Optimizing the Optimizer: Language Models Discover Faster Molecular Relaxation

👍 17

Geometry optimization is a major cost in many quantum-chemical workflows: each optimization step requires one force evaluation, and at the density-functional level that evaluation dominates the wall time. Research in this area has produced a broad range of optimization methods, and we ask whether a

SearchJev: A Fast and Calibrated System-1 Model for Search Agents

👍 14

Search agents repeatedly make short decisions about relevance, evidence sufficiency, and search actions. Using generative language models for these decisions introduces latency and unreliable confidence. We present SearchJev, a fast and calibrated System-1 model that separates search decisions from

Agentic discovery of blood biomarker from distilled private health records

👍 1

Routine complete blood counts (CBCs) could yield new biomarkers, but the private records needed to evaluate candidates cannot be shared with frontier language model agents that excel at discovery. We distilled the evidence held in the Clalit Health Services panel of over 5.4 million patients into a

The Numerical Linear Algebra of Large Language Models

👍 1

Numerical Linear Algebra (NLA) has consistently played a vital role in advancing science by providing tools to solve fundamental problems encountered in scientific and engineering applications. Over the decades, it has continually evolved to meet the demands driven by successive waves of scientific

SEER: Self-Evolving Event Reasoning and Retrieval for Time Series Forecasting

👍 4

Real-world time series are frequently driven by exogenous events and structural shifts, rendering conventional forecasting based solely on historical numerical observations insufficient. While language models can retrieve external news, standard retrieval-augmented approaches struggle with high nois

Learning Latent Protein Languages for Autoregressive Generation

👍 2

Autoregressive transformers remain comparatively weak for protein sequence and structure generation. We study the role of target representation: amino acid tokens encode residue identities without explicit contextual semantics, while backbone coordinates require a discrete representation in our fram

Foresight: planning future perception in streaming VLMs without retraining

👍 0

Existing streaming vision-language models (VLMs) continuously perceive and reason over visual streams, but their computational pathways remain fixed throughout inference. Consequently, they cannot adapt computation to evolving scene dynamics, where different future events demand different levels and

The AI Theorist reveals excitonic structure in α-RuCl_3

👍 0

Advances in experimental instrumentation and automation generate increasingly rich datasets, but turning experimental observations into microscopic understanding remains a bottleneck in scientific discovery. To accelerate this process, we introduce AI Theorist, a system of artificial intelligence (A

World Editing: Intervening on Executable Worlds at Increasing Depth

👍 9

Interactive world models are increasingly capable of generating environments and acting within them, yet deliberately editing an existing executable world remains underexplored. We formulate world editing as intervening on an existing world while preserving properties that should remain unchanged, a

DeskForge: Dense Supervision from Desktop Environments for Computer-Use Agents

👍 7

Computer-use agents need to reliably ground action targets in complex desktop scenes, where multiple applications, overlapping windows, and visually similar controls compete for attention. Existing training data rarely pair such scenes with dense annotations or vary them in a controlled way. We intr

MEA: A Reward-Driven Multi-Agent System for Faithful Model Explanations

👍 0

Recent years have seen the employment of a plethora of machine learning (ML) models in high-stakes domains, but they remain largely opaque to the practitioners who act on their predictions. While post-hoc explanation methods offer a lens into this model behavior, wielding them effectively demands ex

OmniReasoning: Pushing the Limits of Audio-Visual Joint Reasoning

👍 17

Recent advances have enabled unified omni-modal models in understanding audio, vision, and language. However, existing benchmarks, training data, and learning methods largely treat the modalities independently, leaving the capability of audio-visual joint reasoning poorly evaluated and insufficientl

Video2Skill: From Streaming Experience to Reusable Embodied Skills

👍 8

Manipulation behaviors vary widely across objects and scenes, but they share a small set of reusable skills, and planning with these skills helps embodied agents generalize to new tasks. Yet an agent can only plan with skills it knows. Recovering skills from observed experience, the inverse of plann

How To Hire A Grok Bot: Experience From The First Zero Human Company.

@BrianRoemmele · 489.9K 粉丝 · 1.1M 阅 · 533 赞 · 85 转

You do not build a Grok Bot. You hire one. The difference is the whole article. A chatbot answers a question and forgets the room. A Grok Bot has a name, a job, a conversation that persists, and a

中文介绍 探讨如何雇佣Grok Bot,区别于普通聊天机器人,Grok Bot拥有持续对话能力。

He Who Controls the Compute Controls the Universe

@stevenmarkryan · 458.8K 粉丝 · 124.9K 阅 · 541 赞 · 84 转

Incredible things are happening in AI. What’s more incredible is that it’s still EXTREMELY early. As in, Day 0.000. Things are only going to get crazier from here. In case you’ve been living under a

中文介绍 AI发展初期,预测未来将更加疯狂,呼吁关注AI的早期阶段。

Every language learner gets a tutor now

@TryLiveAvatar · 1.1K 粉丝 · 123.4K 阅 · 692 赞 · 188 转

The best way to learn a language was always a patient tutor. For the first time, everyone can have one. Here's what that changes for learners, teachers and schools. The hardest part is speaking Ask

中文介绍 语言学习将迎来变革,AI导师让每个人都能拥有个性化的学习体验。

VIBE MANUFACTURING IS HERE

@gregisenberg · 724.2K 粉丝 · 44.1K 阅 · 569 赞 · 53 转

For the FIRST TIME in HISTORY, you can describe a physical product in one sentence and, a week later, hold it in your hands, cut from metal and built to your exact measurements. That's VIBE

中文介绍 VIBE制造技术突破,首次实现一周内将金属产品设计成实物。

LIST OF HACKATHON FOR OCTOBER.

@onlyonealexia · 1.6K 粉丝 · 26.5K 阅 · 509 赞 · 53 转

A List of Web3, AI & Web2 Hackathons Compiled by @onlyonealexia There are a lot of hackathon lists that look impressive until you actually open the links. I compiled a list of ongoing and upcoming

中文介绍 分享Web3、AI和Web2的Hackathon列表,筛选出值得参与的竞赛。

If you have too many interests, build this AI system in 1 hour

@Hesamation · 90.6K 粉丝 · 21.4K 阅 · 618 赞 · 51 转

if you have an ADHD brain, you probably have six different goals, you get overwhelmed, and you miss days without working on any one of them. that's me. I love agents, harnesses, RL, LLM architecture,

中文介绍 针对多兴趣人士,介绍如何构建AI系统来管理多重目标。

How Jump Trading is scaling quant research with ChatGPT

Jump Trading uses OpenAI to expand quantitative research. See how longer-running AI workflows combine multiple data sources with human review.

中文介绍 Jump Trading利用OpenAI扩展定量研究,通过结合多个数据源和人工审核的AI工作流程。

Sharing AI progress in mathematics

OpenAI publishes new results on open problems in mathematics from an internal frontier model and shares Lean proof formalizations and research details on GitHub.

中文介绍 OpenAI发布关于数学开放问题的内部前沿模型新结果,并在GitHub上分享Lean证明形式化和研究细节。

Advancing computer use with Ironclad

Learn how OpenAI and Ironclad are training and evaluating AI agents on complex contracting workflows to advance computer use for professional work.

中文介绍 OpenAI与Ironclad合作,通过训练和评估AI代理在复杂合同工作流程上的应用,推进计算机在专业工作中的应用。

Connecting AI agents to enterprise knowledge

For all the data that AI systems continually amass and analyze, enterprise AI agents often suffer from a curious shortcoming: a lack of knowledge. More than data, knowledge is the understanding of what the data means in the context of individual organizations. AI agents need this understanding to re

中文介绍 将AI代理连接到企业知识。企业AI代理往往缺乏知识,需要理解数据在特定组织背景中的意义。

Our approach to EU text provenance rules

How OpenAI is approaching text watermarking under EU rules. Learn where watermarks apply, how detection works, and why access starts with researchers.

中文介绍 OpenAI阐述遵守欧盟文本溯源规则的方法,介绍水印应用、检测方式及为何研究者优先获取访问权。

Bringing predictive analytics to the agentic AI era

In 2026, the question for enterprise AI is no longer whether predictive models can outperform statistical forecasts—that argument is settled. The big question now is how to enable predictive systems to act on their own conclusions without drifting from business intent. The frontier has moved from pr

中文介绍 将预测分析引入代理AI时代。企业AI的核心问题是使预测系统能够根据自身结论行动,而不偏离商业意图。

Building advertising for the way people use AI

OpenAI introduces a new visual ad format in ChatGPT and expands measurement tools, attribution partnerships, and brand suitability for advertisers.

中文介绍 OpenAI在ChatGPT中推出新的视觉广告格式,并扩大测量工具、归因合作伙伴关系和品牌适用性。

People really hate AI, so why can’t they get enough?

Over the summer I talked to the CEO of Springboards, a startup building an LLM that’s designed to come up with a wider variety of responses than its mainstream rivals do. At the start of the call, he said something that’s been stuck in my head since: “We often say that we’re a self-loathing AI…

中文介绍 人们真的很讨厌AI,但他们为何不能获得足够的应用?

EmTech Future 2026: When AI Meets Everything

Yossi Matias, Vice President & Head of Google Research, explores how AI is beginning to reshape biology, infrastructure, manufacturing, and science, and why its greatest impact may come when it intersects with other fields. Step inside the newsroom with our MIT Technology Review editors for sharp an

中文介绍 EmTech Future 2026:当AI遇到一切。Yossi Matias探讨了AI如何开始重塑生物、基础设施、制造和科学,以及为什么它最大的影响可能发生在与其他领域的交叉处。

Hunter Valley community group wins landmark high court climate change case

Ruling on Mount Pleasant coalmine shows ‘we cannot continue to dig up coal … and pretend the consequences have nothing to do with us’, group says A Hunter Valley community group has won Australia’s first high court case to consider climate change, in a ruling advocates say sets a binding national pr

Pornhub returns to Australia after banning access due to age verification rules

Australian adults again able to access the world’s most popular porn site after parent company introduces age checking via Apple devices Get our breaking news email, free app or daily news podcast Australian adults will be able to access Pornhub once again after the site brought in age-checking on A

Plague epidemic risk in Russia low after death of lab technician, WHO says

World Health Organization detects no sign of further outbreak since Darya Shipilova died last week The World Health Organization has said the risk of an epidemic in Russia is low after the sudden death of a lab technician who worked at a plague research institute in Siberia. Darya Shipilova, who was

Taiwan Dethrones Korea Atop Global Markets

Taiwan is emerging as the stronger bet for equity investors due to its deeper linkages across the AI supply chain and a more upbeat earnings outlook. Bloomberg's Winnie Hsu breaks down the news. (Source: Bloomberg)

中文摘要 台湾因AI供应链联系更紧密和乐观的盈利预期,成为股民更看好的投资地,超越韩国成为全球市场领导者。

Mary Ng: Asian Trade Growth Can Boost Canadian Exports

Former Canadian Minister of International Trade and Milken Institute Senior Fellow Mary Ng outlines Canada's expanding trade push across Asia and the Indo-Pacific, emphasizing that deeper market access in energy, technology, and agriculture could reach 3 billion consumers and drive export growth. (S

中文摘要 加拿大前国际贸易部长玛丽·吴强调,加拿大在亚洲及印太地区的贸易扩张,有望触及30亿消费者,促进出口增长。

Asset Management One to Boost Talent Spending 30% Over 3 Years

Asset Management One Co. plans to boost spending on talent over the next three years as rising Japanese interest rates broaden investment choices and push pension funds, university endowments and individuals to rethink how they allocate their money.

中文摘要 资产管理公司One计划在未来三年内将人才支出提高30%,以应对日本利率上升带来的投资选择增加。

Jaguar takes a leap into luxury EV market with new model

British carmaker unveils £130,000 electric car following radical rebrand

中文摘要 英国汽车制造商推出了一款价值13万英镑的电动汽车,这是其品牌重塑后的新产品。

Gold Steady as More Oil and Falling Yields Ease Rate-Hike Bets

Gold held gains as increasing oil supplies from the Middle East and a decline in bond yields eased pressure on the US Federal Reserve to hike interest rates this month.

中文摘要 随着中东地区石油供应增加和债券收益率下降,黄金价格保持稳定,缓解了美联储本月加息的压力。

BHP Sells Shuttered Australian Nickel Plant to Gold Fields

BHP Group will sell its Kambalda nickel concentrator plant and associated land in Western Australia to Gold Fields Ltd. for an undisclosed price.

中文摘要 必和必拓公司将位于澳大利亚的Kambalda镍精炼厂及其相关土地出售给Gold Fields Ltd.,交易价格未公开。

Samsung Investors Want Evidence Profit Boom Has Staying Power

Samsung Electronics Co.’s earnings will test whether it can convince investors of its long-term outlook, an increasingly critical task as record profits have failed to reinvigorate the stock.

中文摘要 三星电子的收益将考验其能否让投资者相信其长期前景,因为创纪录的利润未能提振股价。

Ken Leech Agrees to $3 Million SEC Fine Over Cherry-Picking Case

Ken Leech, the former co-chief investment officer at Western Asset Management Co., agreed to pay $3 million to settle a US Securities and Exchange Commission lawsuit claiming he’d cherry picked winning trades.

中文摘要 西部资产管理公司前联合首席投资官肯·利奇同意支付300万美元以解决美国证券交易委员会对其选择性交易的诉讼。

Froyo's made a comeback. But at £12 a tub will it last?

Frozen yoghurt, which was a huge craze in the 2000s and 2010s, has made its return with multiple chains popping up across the wider UK.

中文摘要 冰沙卷土重来,但每桶12英镑的价格能维持多久?

该源今日无内容。