Global Research Tracker

全球顶尖实验室科研动态

追踪 MIT · Stanford · Berkeley · CMU · DeepMind 等全球顶尖高校与 AI 机构的最新研究方向、论文成果与实验室动态
🕐 最近更新 2026-08-04 19:37
📡 RSS 动态 44 条
📄 arXiv 论文 80 篇
🇨🇳 国内机构 5 家
🌍 全球机构 10 家
Research Directions

顶尖实验室研究方向

Global Top Lab Directory

精选全球 AI / CS 领域前沿实验室,覆盖大模型、高效推理、具身智能、AI4Science 等核心方向

💬

大语言模型与 Agent热门

LLM & AI Agents
Stanford NLP Group大模型对齐、知识推理、对话系统
Princeton NLPLLM 评估、安全性、语言理解生成
Berkeley BAIR强化学习、多模态 Agent、规划
CMU Language Technology多语言、低资源、对话系统

高效推理与模型压缩核心

Efficient Inference & Compression
MIT Han Lab量化、剪枝、蒸馏、神经架构搜索 (NAS)
Stanford MLSysML 系统优化、分布式训练、推理加速
Berkeley Sky Lab云际计算、跨云 AI 调度
🔬

AI for Science前沿

AI for Science
Google DeepMind Science蛋白质结构预测、数学推理、气候模型
Stanford AIMI医学影像、临床决策支持
MIT CSAIL基因组学、药物发现、生物计算
🦾

具身智能与机器人前沿

Embodied AI & Robotics
Berkeley RLL机器人学习、抓取与操作
CMU Robotics Institute自主移动、人机协作、感知
Stanford Robotics灵巧手操作、强化学习
🖥️

AI 系统与算力调度核心

AI Systems & HPC
Berkeley Sky Computing跨云调度、算力互联、低延迟
Stanford DAWNAI-数据库融合、高效系统
CMU PDL分布式存储、数据中心体系结构
🌱

绿色 AI 与可持续计算方向

Green AI & Sustainable Computing
MIT Energy Initiative能源效率优化、绿色数据中心设计
ETH Zurich Computing低功耗 AI 芯片、算法-硬件协同设计
🏛️

中国顶尖高校 AI 实验室国内

China's Top University Labs
清华大学 AIR面向产业的 AI 研究、大模型、具身智能
清华大学 THMLNLP、知识图谱、推理与对话
北京大学 AI 研究院视觉-语言多模态、智能系统
上海交大 APEX Lab数据挖掘、推荐系统、LLM 应用
浙江大学 CAD&CG 国重计算机图形、3D 视觉、数字人
复旦大学 FudanNLPNLP、语义理解、医疗 AI
🔭

中国科研机构与 AI 实验室国内

China's Research Institutes & AI Labs
上海人工智能实验室书生·浦语 InternLM · 多模态 · 开源生态
智源研究院 BAAI悟道大模型 · FlagAI · AI 评测基准
之江实验室脑机接口 · 人机协同 · 超算基础设施
中科院自动化所模式识别 · 计算机视觉 · 脑科学
鹏城实验室鹏城云脑 · 开源 AI 基础设施
🚀

中国头部 AI 企业研究院国内

China's Top AI Company Labs
DeepSeek 深度求索DeepSeek-R1/V3 · MoE 架构 · 低成本高效训练
智谱 AIGLM 系列 · ChatGLM · Agent · 视频生成
百度研究院文心大模型 · 飞桨 PaddlePaddle · 自动驾驶
阿里达摩院通义千问 Qwen · 多模态 · 云端 AI
华为诺亚方舟实验室盘古大模型 · 昇腾 NPU · 联邦学习
字节跳动 SeedDoubao · 视频生成 · 多模态理解
月之暗面 Moonshot AIKimi · 超长上下文 · Reasoning
MiniMax语音合成 · 视频生成 · Agent 框架
🇨🇳 China AI Updates

中国境内 AI 动态

Domestic Research Institutes · Enterprise Labs · Media

覆盖上海 AI Lab · 智源 BAAI · DeepSeek · 智谱 · 百度 · 达摩院 · 全球计算联盟 · 机器之心 · 量子位等机构与媒体

🇨🇳
雷峰网 雷峰网 AI 媒体
AI 硬件 · 大模型落地 · 智能驾驶 · 算力
Tue, 04 Au
超400万人在灵光App“手搓”AI应用,加速AI原生创作者生态形成
AI应用创作正从专业开发走向普罗大众。8月3日,灵光App宣布,其平台“闪应用”创作者已超过400万人,其中绝大多数为没有编程背景的普通用户,包括学生、教师、心理疗愈师、游戏发烧友、网文爱好者等群体。随着创作者规模不断扩大、应用类型日趋丰富,一个由普通用户驱动的AI原生应用生态正在加速形成。平台数据显示,游戏化是当前最活跃的趋势。其中,“模拟器”成为今年上半年最热门类型,诞生近万个相关应用,覆盖升学、职场、偶像养成等主题。与此同时,AI
Tue, 04 Au
算力筑基 智云赋能 | 仪电智算全力护航全国青少年人工智能大赛决赛
8月1日至8月3日,由中国福利会和中国妇女发展基金会共同主办的首届全国青少年人工智能大赛全国决赛在上海举办。作为首个落户上海的自然科学素养类白名单赛事,大赛以“智能向善 生长无限”为主题,设立五大前沿赛道,吸引全国近4万名中小学生参与初赛,数百名选手齐聚上海里滩文化中心角逐桂冠。作为中国福利会的战略合作伙伴,上海仪电携旗下企业智算服务为人工智能驱动科学、具身智能、大语言模型应用三大决赛赛道提供高性能算力服务与仪电智算云YiCloud平台
Tue, 04 Au
参数内卷的尽头,泳池机器人在等待一次范式转移
销量、品牌和渠道座次正在迅速成形,但泳池机器人的真正技术分水岭才刚刚出现。过去几年,无线化降低了产品的使用门槛,也打开了亚马逊、大型商超和泳池专业渠道。率先抓住这一窗口的公司,已经建立起销量、品牌和货架优势。后来者面对的,不只是产品差距,还有漫长的渠道验证周期。摆脱电线之后,许多泳池机器人仍然依赖反复折返完成清洁。面对异形池底、台阶、浅水区和水线,它们很难准确判断自己在哪里、哪些区域已经清洁,漏扫、重复运行和人工干预仍然普遍存在。这构成
Tue, 04 Au
发布当日,海外主流AI平台纷纷接入阿里Qwen3.8
8月3日,阿里巴巴正式发布新一代基座大模型Qwen3.8,编程与Agent能力大幅提升,整体性能位居全球大模型第一梯队。发布当日,OpenRouter、OpenCode、Hermes Agent、Command Code、Vercel、Novita、Charm、DeepInfra等多家海外主流API聚合、Agent工具及开发者平台纷纷拥抱Qwen3.8-Max,形成了一股“接入潮”。Hermes Agent、OpenCode、Comma
🇨🇳
全球计算联盟 GCC 全球计算联盟
智算基础设施 · 产业标准 · 全球计算生态
2026-07-30
GCC加入openEuler香港用户组,软硬协同共建亚太开源算力生态
作为中国首个计算领域的国际性产业与标准组织,全球计算联盟(GCC)在成立短短两年中加速推进国际化布局,相继与多家欧亚伙伴建立了立体化合作关系。继东南亚Joint Hub成立后,GCC近日又迈出关键一步——加入openEuler香港用户组,以软硬协同能力赋能香港计算产业发展。
2026-07-20
《超节点定义与实践白皮书》正式发布!确立AI基础设施建设新范式
【中国・上海,2026年7月19日】在2026世界人工智能大会(WAIC)期间,鹏城实验室(PCL)主任高文和全球计算联盟(GCC)理事长金海共同发布《超节点定义与实践白皮书》。本白皮书由GCC和PCL联合牵头编制,首次完整厘清超节点核心定义、架构特征与核心技术体系,明确三大技术特征、梳理十大核心价值场景、汇集全产业链标杆落地案例,为全球下一代智能计算基础设施建设划定技术标准、指明演进方向、提供实践范本。
2026-06-30
GCC月览|新会员涌入、HiFloat首秀欧洲、OAII社区半年会将启…GCC六月全速推进
六月,万物峥嵘,热潮涌动! 14家新成员加入GCC大家庭,覆盖光互连、智算芯片、工业自动化、检测认证等多个技术赛道;Open AI Infra社区年中论坛亮点与完整版议程重磅发布;HiFloat首秀欧洲ISC 2026,获得海内外技术专家广泛关注;更有NPO光互连、具身智能、机密计算专题活动圆满收官……GCC及各分支机构持续面向全球产业伙伴开放,欢迎了解、加入!
2026-06-26
CEIC2026 筹备会在深召开:创新企业代表一起定义消费电子新风向
6月24日,2026消费电子创新大会(CEIC2026)筹备会在深圳市民中心召开。本次会议集结通信、半导体、具身智能、数字健康、低空经济、超高清视听等前沿赛道的创新企业及国际标准与产业组织代表,围绕产业趋势预判、新技术生态落地、分论坛体系搭建、特色活动设计四大核心议题展开深度研讨。与会各方就人工智能全面驱动下的消费电子行业新趋势达成统一行动共识,企业代表结合各自领域技术商业化实践分享真知灼见,为CEIC这一年度创新盛会的顶层策划与内容落地提供了实操依据。
🌍 Global Lab Updates

全球实验室动态

Live RSS · BAIR · MIT · Stanford · DeepMind · OpenAI · NVIDIA

来自全球顶尖 AI 机构官方博客与新闻的最新内容,自动抓取更新

arXiv · Latest Papers

顶尖机构近期论文

Recent Papers from Elite Institutions

从 arXiv 抓取 cs.LG / cs.AI / cs.CL / cs.CV / cs.RO 分类中顶尖机构的最新论文

人工智能 · AI
2026-08-04
Computing Actual Causes for Neural Network Predictions under Structured Causal Inputs
Explaining the predictions of neural networks is a central challenge in trustworthy AI. Existing explanation methods, such as those based on feature attribution or minimal sufficient sets, typically treat input features as independent, which can yield misleading explanations when inputs exhibit stru…
2026-08-04
MDLMPE: Distribution Aware Positional Encoding for Masked Diffusion Language Models
Masked diffusion language models (MDLMs) enable parallel generation and bidirectional context modeling, but their positional context differs fundamentally from that of autoregressive (AR) models. Whereas AR decoding exposes a contiguous prefix, MDLM denoising produces dynamic, non-contiguous configu…
2026-08-04
GDPevo: Evaluating Agent Self-Evolution on Real Business Tasks
Agent self-evolution updates an agent's persistent state from prior experience and reuses it to solve related tasks more effectively. Evaluating self-evolution is difficult: existing benchmarks provide limited coverage of economically valuable task domains, do not always design training and test tas…
2026-08-04
Risky Business: Measuring The Faithfulness-Safety Tension
Chain-of-Thought (CoT) reasoning offers a promising window into model monitoring. However, monitoring relies on faithfulness, i.e., the model output strictly derives from its reasoning trace. We identify an alignment tension where a model must be faithful enough to be monitored, yet robust enough to…
2026-08-04
Agents Catching Agents: Shortcut Cascades and Benchmark Gaming in Clinical Multi-Agent Systems
Clinical decision support is moving toward committees of language-model agents deliberating on a shared workspace. We ask whether such committees can be gamed by shortcuts, cues a benchmark rewards but a clinician would ignore. Across seven cohorts on six public datasets spanning text (MedQA-USMLE, …
自然语言处理 · NLP
2026-08-04
MDLMPE: Distribution Aware Positional Encoding for Masked Diffusion Language Models
Masked diffusion language models (MDLMs) enable parallel generation and bidirectional context modeling, but their positional context differs fundamentally from that of autoregressive (AR) models. Whereas AR decoding exposes a contiguous prefix, MDLM denoising produces dynamic, non-contiguous configu…
2026-08-04
Risky Business: Measuring The Faithfulness-Safety Tension
Chain-of-Thought (CoT) reasoning offers a promising window into model monitoring. However, monitoring relies on faithfulness, i.e., the model output strictly derives from its reasoning trace. We identify an alignment tension where a model must be faithful enough to be monitored, yet robust enough to…
2026-08-04
An Actionable Diagnosis of Multilingual, Multi-Agent Planning Failures
Multilingual multi-agent systems exhibit substantial degradation beyond English, yet prior work rarely identifies how task-critical information is lost when user requests are converted into executable plans. We study the planner in a multi-agent system as the request-to-action interface and derive a…
2026-08-04
GPTKB 2.0: Direct Construction of Disambiguated Knowledge Bases from Large Language Models
Automated Knowledge Base Construction (AKBC) is a core NLP task, and recent work proposes generating knowledge bases directly from large language models (LLMs), treating the model itself as the knowledge source. However, LLMs natively possess no representation of entities, leading to duplicate entri…
2026-08-04
When Outputs Disperse, Does Epistemic Revision Follow? A Black-Box Coupling Diagnostic for Machine Collectives
Collective intelligence research treats disagreement as evidence of epistemic diversity: if agents express different views, the group should retain capacity to revise. In LLM collectives this proxy can break: agents can produce diverse-looking arguments while preserving the same conclusion. We opera…
计算机视觉 · CV
2026-08-04
TDVR: Joint Text Disambiguation and Viewpoint Reasoning for Zero-Shot 3D Visual Grounding
Zero-shot 3D visual grounding aims to localize specific objects based on textual descriptions and 3D visual input. However, the effectiveness of existing methods is significantly hindered by the ambiguous query text and deficient viewpoints. To address these issues, we propose TDVR, a training-free …
2026-08-04
Unsupervised Adversarial Domain Adaptation for Uterine layer Segmentation: From Labeled Cine to Unlabeled Dynamic EPI MRI
Uterine peristalsis is a key physiological phenomenon responsible for various functions across the menstrual cycle, intimately linked to uterine wall microstructure. Alterations in uterine motion and tissue properties are implicated in the etiology of gynecological diseases, yet these processes have…
2026-08-04
Towards Reliable and Reproducible Fetal Brain Biometry: A Deep Learning Approach Using MRI
Fetal brain biometry is essential for quantitative assessment of brain development, supporting gestational age estimation, developmental monitoring, and detection of abnormalities. In clinical practice, measurements are manually performed, making them time-consuming and prone to variability. While a…
2026-08-04
Attention is Case-Sensitive
In human visual perception, uppercase lettering serves as a natural salience cue that captures attention within lowercase text. In this paper, we present a systematic empirical characterization study revealing that Large Language Models (LLMs) exhibit an analogous property: letter casing modulates i…
2026-08-04
MultiCompose: Multi-Concept Personalized Composition with Per-Subject Attribute Binding
Text-to-image diffusion models enable personalization of specific visual concepts from a small number of reference images. However, generating a single image that contains multiple personalized subjects, each bound to user-specified attributes such as clothing, accessories, and held objects, remains…
机器人学 · Robotics
2026-08-04
GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation
Learning long-horizon manipulation skills with reinforcement learning remains challenging due to the complexity of reward design, the limited guidance of sparse rewards, and the high cost of manual subtask annotation. Visual demonstrations can provide supervision for reward learning, but rewards lea…
2026-08-04
Track4Action: Distilling World-Centric 3D Tracker into Vision-Language-Action Policies
Action labels tell a vision-language-action (VLA) policy which robot commands to imitate, but not how those commands change the 3D world. The aligned demonstration clip contains this missing supervision because its $K$ frame transitions record the geometry, motion, visibility, and camera change prod…
2026-08-04
LiLa-WAM: Lightweight Latent Reasoning World-Action Model for Robotic Manipulation
World-action modeling has emerged as a promising paradigm for robotic control, as it empowers models to go beyond reacting to observations and anticipate how a scene will evolve. However, existing WAMs often incur substantial computational overhead. Pixel-space methods often allocate substantial cap…
2026-08-04
PhyAI: Real-Time Physical AI at the Edge, Scalable Rollouts in the Cloud
Physical AI policies require inference throughout their lifecycle, including model evaluation, cloud reinforcement learning rollout, edge GPU serving, and onboard deployment. Although these settings share the same checkpoint and action semantics, they often rely on separate inference programs. To un…
2026-08-04
Active Stiffness Control of a Supportive Continuum Robot
Supportive continuum robots (SCRs) enhance the load-bearing capability of an operative continuum robot by mechanically coupling it with a supportive arm. However, their passive stiffness is determined by the mechanical configuration and cannot be adjusted online for varying payloads or interaction f…