yl-1993/learn-to-cluster
Learning to Cluster Faces (CVPR 2019, CVPR 2020)
Learning to Cluster Faces (CVPR 2019, CVPR 2020)
Vision as Unified Multimodal Generation
SenseNova-U series: Native Unified Paradigm with NEO 2.0 from the First Principles
A PyTorch Implementation of Matrix Capsules with EM Routing
In our implementation of Qwen-Image-Edit, we employ block causal attention to improve inference speed.
Scaling Spatial Intelligence with Multimodal Foundation Models
Holistic Evaluation of Multimodal LLMs on Spatial Intelligence
Accelerated Training for Massive Classification via Dynamic Class Selection (AAAI 2018, Oral)
An online GUI tool used to visualize prototxt and generate prototxt for caffe (current version).
Spectral Method for Multiple Experts Inverse Reinforcement Learning
Fastfood transform layer for caffe
OpenXRLab Human Motion Generation Codebase
OpenMMLab Human Pose and Shape Estimation Toolbox and Benchmark
OpenXRLab Synthetic Data Rendering Toolbox
Open3D: A Modern Library for 3D Data Processing
MANO layer for PyTorch, generating hand meshes as a differentiable layer
OpenXRLab foundational library for XR-related algorithms
OpenXRLab Multi-view Motion Capture Toolbox and Benchmark
A PyTorch Implementation of ConvDeltaOrthogonal Initializer
Implementation of the Zhang-Suen thinning algorithm using OpenCV
Matrix Multiple under MapReduce
OpenMMLab Semantic Segmentation Toolbox and Benchmark.
OpenMMLab Detection Toolbox and Benchmark
Open MMLab Computer Vision Foundation
Computation using data flow graphs for scalable machine learning
Music composition based on image feature
Scalable Machine Learning with Dask