GitHub GrowthRelevance 90
aipoch/open-science: +95 GitHub stars
Open Science is an open-source, local-first, model-agnostic AI research workbench for scientific discovery.
Open sourceGitHub GrowthRelevance 90
synthetic-sciences/openscience: +22 GitHub stars
The open-source AI workbench for scientific research
Open source36KrRelevance 90
AI的下一个 Claude Code,可能诞生在实验室
过去两年,AI率先在Coding领域跑通了高价值商业闭环。 海外机构预计,未来12个月,AI Coding有望为OpenAI和Anthropic带来合计约1200亿美元的年化收入。Coding是推动其商业化增长的重要场景之一。 Coding之所��成为大模型最早兑现价值的专业场景,关键并不只是程序员成本高,而是代码天然拥有编译、测试和运行结果这套“验证器”:模型每生成一段代码,都能迅速知道对错,并将能力提升直接转化为生产力。 但随着Coding逐渐成为市场共识,投资人也开始寻找AI商业化的下一个高价值入口。 科学研...
Open sourceOpenAIRelevance 90
Ten advances in mathematics and theoretical computer science
OpenAI shares new results on long-standing open problems in mathematics and theoretical computer science, including advances in geometry, cryptography, and complexity.
Open sourcex manual globalRelevance 90
We’re giving scientists, mathematicians, and engineers free access to our frontier models—starting with 10,000 researchers and expanding to 100,000 through 2027.
We’re giving scientists, mathematicians, and engineers free access to our frontier models—starting with 10,000 researchers and expanding to 100,000 through 2027. ChatGPT for Academic Researchers is built to accelerate discovery across disciplines.
Open sourceRedditRelevance 90
[Paper] SWE-Pruner Pro: The Coder LLM Already Knows What to Prune
Pruning long context for coding agents has been a vital technology for efficient context management. While existing context pruning methods such as SWE-Pruner realize this by attaching a separate code classifier, we find the agent itself encodes internal repre...
Open sourceGitHub GrowthRelevance 90
synthetic-sciences/openscience: +35 GitHub stars
The open-source AI workbench for scientific research
Open sourceHF Daily PapersRelevance 90
DeepSearch-World: Self-Distillation for Deep Search Agents in a Verifiable Environment
Training tool-use agents to improve from their own experience remains challenging, as supervised fine-tuning relies on fixed teacher-distilled trajectories, while sparse-reward reinforcement learning provides weak supervision for long-horizon interactions. We ...
Open sourceHF Daily PapersRelevance 90
SWE-Pruner Pro: The Coder LLM Already Knows What to Prune
Pruning long context for coding agents has been a vital technology for efficient context management. While existing context pruning methods such as SWE-Pruner realize this by attaching a separate code classifier, we find the agent itself encodes internal repre...
Open sourceHF Daily PapersRelevance 90
ReflectWorld-MM: An Entity-Oriented Multimodal Memory System for Open-Ended Video Streams
Building assistants that can continually watch the world, remember what they see, and reason over their accumulated experience is a long-standing goal, and recently multimodal agents equipped with long-term memory over video streams have attracted increasing i...
Open sourceHF Daily PapersRelevance 90
JoyNexus: Service-Oriented Multi-Tenant Post-Training for VLA Models
The post-training of Vision-Language-Action (VLA) models is essential due to the diversity of simulators, robot embodiments, and task objectives. Existing compute services, whether offered as direct accelerator rental or batch-workload submission, typically al...
Open sourcehnRelevance 90
Claude Fable produced a counterexample to the Jacobian Conjecture
hello there the jacobian conjecture is false thanx to my close friend akhil for asking about it and my other close friend fable for working during the world cup final ((1+xy)^3 z + y^2 (1+xy) (4+3xy), y + 3 x (1+xy)^2 z + 3 x y^2 (4+3xy), 2 x - 3 x^2 y - x^3 z...
Open sourceHF Daily PapersRelevance 90
DSWorld: A Data Science World Model for Efficient Autonomous Agents
Despite strong capabilities in data understanding and decision-making, autonomous data science agents still heavily rely on trial-and-error workflows that involve expensive computation. This bottleneck motivates models that can anticipate the effects of data s...
Open sourceHF Daily PapersRelevance 90
RESOURCE2SKILL: Distilling Executable Agent Skills from Human-Created Multimodal Resources
Skills are a useful abstraction for software agents, turning human and agent experience into reusable procedural knowledge. Yet existing skill libraries are mostly hand-written, text-centric, or derived from agent traces, leaving tutorial videos and other mult...
Open sourceHF Daily PapersRelevance 90
Recursive Harness Self-Improvement
Under model--harness co-evolution, harnesses are not merely inference-time scaffolds but data-generating components whose execution traces can shape future foundation models. This motivates harness-in-the-loop learning: optimizing harnesses for both immediate ...
Open sourceHF Daily PapersRelevance 90
From Human-Centric to Agentic Code Review: The Impact of Different Generations of Generative AI Technology on Review Quality
Code review helps maintain software quality before code integration, but it also imposes a substantial workload on human reviewers. As generative artificial intelligence becomes part of software development, code review is shifting from a primarily human revie...
Open sourcehnRelevance 90
Solving 20 Erdős Problems with 20 Codex Accounts Running in Parallel
A Mac app controlling up to 20 starship agentic harnesses running GPT-5.6 Sol max in parallel on open mathematics problems. 19 Erdős problems 0 Frontier Math problems 0 Millennium problems Star Fleet Math Built by Colin Snyder · colin@colinsnyder.com Advised b...
Open sourceHF Daily PapersRelevance 90
Towards Autonomous and Auditable Medical Imaging Model Development
Large language model (LLM) agents are beginning to automate machine learning engineering (MLE) by coupling planning, code execution, debugging, and empirical feedback. Translating this capability to medical imaging remains difficult because each task imposes m...
Open sourcebluesky globalRelevance 90
Anthropic’s Claude Science is coming for Kendall Square
AI still has the potential to do more than reduce corporate head counts. Watch out, Kendall Square. www.statnews.com/2026/07/14/a... AI still has the potential to do more than reduce corporate head counts. Watch out, Kendall Square.
Open sourceTechCrunchRelevance 90
OpenAI researcher Miles Wang in talks to launch AI drug discovery startup valued at $2B
The funding discussions point to investor interest in applying AI to make breakthroughs in life sciences.
Open sourceNVIDIARelevance 90
How to Run an Autoresearch Workflow with RL Agent Skills and NVIDIA NeMo
Coding AI agents are becoming practical operators for long-running machine learning (ML) workflows. They can inspect repositories, set up runtimes, resolve...
Open sourceNVIDIARelevance 90
Post-Train NVIDIA Cosmos 3 in One Day Using Agent Skills
What if autonomous coding AI agents could push your vision reasoning models above 90% accuracy with almost no manual effort? When adapting vision reasoning...
Open source