公司库

Google DeepMind

全球顶尖AI研究实验室;AlphaGo/AlphaFold/RT-2等里程碑成果;机器人学习与具身智能前沿;Gemini多模态模型赋能机器人;被Google/Alphabet收购

伦敦 2010 年成立 模型/算法公司 / 通用AI与机器人学习 上市

公司介绍

全球顶尖AI研究实验室;AlphaGo/AlphaFold/RT-2等里程碑成果;机器人学习与具身智能前沿;Gemini多模态模型赋能机器人;被Google/Alphabet收购

相关人物

相关模型

相关论文

WaveNet: A Generative Model for Raw Audio

原始音频波形生成的因果扩张卷积模型,奠定神经声码器方向。

Grandmaster Level in StarCraft II Using Multi-agent Reinforcement Learning

在星际争霸II击败人类职业选手的多智能体强化学习系统。

A General Reinforcement Learning Algorithm that Masters Chess, Shogi, and Go Through Self-Play

通用自博弈强化学习算法,在围棋/国际象棋/将棋超越专用AI。

Grammar as a Foreign Language

用序列到序列模型做句法分析,把语法解析当作翻译任务。

Continual Learning with Deep Networks: A Review and Perspective

深度网络持续学习综述:方法、评估与开放问题。

Neural Programmer-Interpreters

用神经网络学习可解释程序的递归执行框架。

MapReduce: Simplified Data Processing on Large Clusters

MapReduce分布式数据处理,大数据系统奠基。

Bigtable: A Distributed Storage System for Structured Data

Bigtable分布式存储,NoSQL系统奠基。

Sequence to Sequence Learning with Neural Networks

Seq2Seq序列到序列学习,NLP里程碑。

Neural Architecture Search with Reinforcement Learning

NASNet神经架构搜索,用RL搜索网络结构。

Distilling the Knowledge in a Neural Network

知识蒸馏,模型压缩与教师-学生范式奠基。

Retrieval-Augmented Language Model Pre-Training

REALM检索增强语言模型预训练,RAG前身。

Understanding Deep Learning Requires Rethinking Generalization

重新思考泛化,深度学习理论里程碑。

Building High-level Features Using Large-scale Unsupervised Learning

Google Brain大规模无监督学习发现高层特征(猫脸神经元)。

Safe Reinforcement Learning by Bootstrapping

安全RL综述,Kohli可信AI方向。

RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control

RT-2 把视觉语言模型迁移到机器人控制,实现 VLA 统一架构。

Mastering Diverse Domains through World Models

DreamerV3 用世界模型在 150+ 任务上单一超参数达到 SOTA,证明世界模型可行性。

Mastering the game of Go with deep neural networks and tree search

AlphaGo 用深度神经网络与蒙特卡洛树搜索击败围棋世界冠军李世石。

Highly accurate protein structure prediction with AlphaFold

AlphaFold 通过深度学习实现高精度蛋白质结构预测,解决生物学五十年难题。

Training Compute-Optimal Large Language Models

提出 Chinchilla 缩放定律:数据与算力需同步增长,70B 模型可达更优性能。

Human-level control through deep reinforcement learning

DQN 用深度网络学习 Q 值,在 Atari 游戏上达到人类水平。

产业情报(参考)

以下字段来自第三方产业图谱整理,仅供参考,不作为投资依据。

六维评分
  • 技术 20
  • 创新 68
  • 资本 40
  • 产品 36
  • 热度 19
  • 组织 24