Articles by Dmytro Spodarets
163 items

AutoAvatar: Autoregressive Neural Fields for Dynamic Avatar Modeling
In their new work, AutoAvatar, researchers have made implicit avatar modeling possible for the first time. It is an autoregressive approach for modeling dynamically deforming human bodies directly from raw scans. 0:00/1× Animated 3D models of the human body are a key tool for applications ranging from virtual fitting to social telepresence. AutoAvatar models body geometry implicitly -- using a signed distance field (SDF) -- and can learn directly from raw scans without requiring temporal match
Nov 1, 2022
High Fidelity Neural Audio Compression
Meta Fundamental AI Research (FAIR) team on audio hypercompression shows how AI can be used to ensure that audio messages don't glitch or slow down when the Internet connection is poor. AI researchers have created a three-part system and trained it to compress audio data to a given size. This data could then be decoded using a neural network. They achieved about 10 times the compression rate of MP3 at 64 kbps without loss of quality and were the first to apply it to 48 kHz stereo audio (i.e. CD
Oct 31, 2022
Data Phoenix Digest - ISSUE 56
Webinar "How we built a recommendation system from scratch", reinforcement learning with SARSA, how I passed the AWS ML Specialty Certification, language understanding with BERT, LION, Omni3D, EVA3D, Text2Light, Modelverse, news, courses, and more.
Oct 28, 2022
Prompt-to-Prompt Image Editing with Cross-Attention Control
Large-scale text-driven fusion diffusion models have attracted a lot of attention because of their remarkable ability to generate a wide variety of images that follow given text cues. Based on these fusion models, it has become natural to create text-driven image editing capabilities. But because it is an inherent property of editing techniques to retain some of the content of the original image, whereas in text-based models, even a small change to a textual cue often leads to a completely diffe
Oct 26, 2022
LION: Latent point diffusion models for generating 3D shapes
Diffusion denoising models (DDM) have shown promising results in the synthesis of 3D point clouds. For 3D DDM to get better, it requires high quality generation, flexibility for manipulation and applications such as conditional synthesis and shape interpolation, and the ability to output smooth surfaces or meshes. At the NeurIPS 2022 conference, a novelty in the AI world is presented, the Latent Point Diffusion Model (LION). This is a 3D shape generation DDM that focuses on training a 3D genera
Oct 26, 2022
Omni3D: A large reference and model for detecting 3D objects in the wild
Recognizing scenes and objects in 3D from a single image is a long-standing goal of computer vision, with applications in robotics and AR/VR. After the success of 2D recognition, Meta AI returns to the task of detecting 3D objects by introducing a large benchmark called Omni3D that uses and merges existing datasets, resulting in 234,000 images annotated with more than 3 million instances and 97 categories. The new Cube R-CNN, trained on Omni3D, is designed to summarize all camera types and scen
Oct 24, 2022
Data Phoenix Digest - ISSUE 55
Charity AI webinar about synthetic data, introduction to DVC and MLflow for experiment tracking, vision transformer model, neural density-distance fields, 40 open-source audio datasets for ML, MDM, DiffDock, YOLO-FaceV2, news, tools, and more.
Oct 18, 2022
Data Phoenix Digest - ISSUE 54
Charity webinar "The promising role of synthetic data to enable responsible innovation", YOLOV7 object counter logic, AutoML for object detection, neural density-distance fields, retrieval-augmented diffusion models, MinVIS, CLIFF, OFA, news, courses, tools, and more.
Oct 10, 2022
Data Phoenix Digest - ISSUE 53
Charity AI webinar, review YOLOv6, GNN in TensorFlow, recent advances and applications of DL methods in materials science, a conversational paradigm for program synthesis, YOLOv7, ASE, PyMAF-X, CVPR 2022 papers, datasets, news, courses, and more.
Sep 13, 2022
Data Phoenix Digest - ISSUE 52
DALL-E is now available in beta, and DALL-E 2 prompt book, introduction to diffusion models for ML, distribute your PyTorch model in less than 20 lines of code, k-means mask transformer, explaining chest X-ray pathologies in natural language, FIGS, NU-Wave, VQAD, datasets, tools, courses, and more.
Aug 17, 2022
Data Phoenix Digest - ISSUE 51
YOLOv7, from ML model to ML pipeline, MLOps and ML roadmap, training the YOLOv5, multiplying matrices without multiplying, SoundSpaces platform, the Shapley value in ML, MineDojo, OmniBenchmark, HaGRID, PMData, courses, and more.
Aug 1, 2022
Data Phoenix Digest - ISSUE 50
Machine Learning & Data Science Survey 2022, how to test ML models in the real world, best practices for deploying language models, neural 3D reconstruction in the wild, mask DINO, MotionCNN, CVNets, StylizedNeRF, courses, tools, and more.
Jul 20, 2022