Bringing the world closer together by advancing AI

Deepfake Detection

Bringing the world closer together by advancing AI

DensePose

Bringing the world closer together by advancing AI

Detectron2

Bringing the world closer together by advancing AI

Deepfake Detection

Bringing the world closer together by advancing AI

DensePose

Bringing the world closer together by advancing AI

Detectron2

Open-Source AI Tools

We share our open source frameworks, tools, libraries, and models for everything from research exploration to large-scale production deployment.

Open-Source AI Research

We're advancing the state-of-the-art in artificial intelligence through fundamental and applied research in open collaboration with the community.

Notable Papers

COMPUTER VISION

Live Face De-Identification in Video

Oran Gafni

Lior Wolf

Yaniv Taigman

International Conference on Computer Vision (ICCV)

RESEARCH

COMPUTER VISION

TensorMask: A Foundation for Dense Object Segmentation

Xinlei Chen

Ross Girshick

Kaiming He

Piotr Dollar

International Conference on Computer Vision (ICCV)

RESEARCH

Single-Network Whole-Body Pose Estimation

Gines Hidalgo

Yaadhav Raaj

Haroon Idrees

Donglai Xiang...

International Conference on Computer Vision (ICCV)

SPEECH & AUDIO

A Universal Music Translation Network

Noam Mor

Lior Wolf

Adam Polyak

Yaniv Taigman

International Conference on Learning Representations (ICLR)

Latest Publications

Asynchronous Gradient-Push | Facebook AI Research

We consider a multi-agent framework for distributed optimization where each agent has access to a local smooth strongly convex function, and the collective goal is to achieve consensus on the parameters that minimize the sum of the agents’…

Mahmoud Assran, Michael Rabbat

GrokNet: Unified Computer Vision Model Trunk and Embeddings For Commerce

In this paper we propose image classification modeling technique targeted for marketplace. We use public posts from marketplace and search log interactions for training image classifier and achieve significant improvements in e-commerce in comparison to previous version of our image classifier.

Sean Bell, Yiqun Liu, Sami Alsheikh, Yina Tang, Ed Pizzi, M. Henning, Karun Singh, Omkar Parkhi, Fedor Borisyuk

COMPUTER VISION

PIFuHD: Multi-Level Pixel-Aligned Implicit Function for High-Resolution 3D Human Digitization

Due to memory limitations in current hardware, previous approaches tend to take low resolution images as input to cover large spatial context, and produce less precise (or low resolution) 3D estimates as a result. We address this limitation by formulating a multi-level architecture that is end-to-end trainable

Shunsuke Saito, Tomas Simon, Jason Saragih, Hanbyul Joo

Iterative Answer Prediction with Pointer-Augmented Multimodal Transformers for TextVQA | Facebook AI Research

Many visual scenes contain text that carries crucial information, and it is thus essential to understand text in images for downstream reasoning tasks. For example, a deep water label on a warning sign warns people about the danger in the…

Ronghang Hu, Amanpreet Singh, Trevor Darrell, Marcus Rohrbach

Help Us Pioneer the Future of AI

We share our open source frameworks, tools, libraries, and models for everything from research exploration to large-scale production deployment.