13 66 6

Ming Li

limingcv

https://liming-ai.github.io

liming-ai

AI & ML interests

Computer Vision, AIGC, VLM/LLM

Recent Activity

updated a dataset 4 days ago

limingcv/cc

published a dataset 4 days ago

limingcv/cc

authored a paper 5 days ago

How Instruction and Reasoning Data shape Post-Training: Data Quality through the Lens of Layer-wise Gradients

View all activity

Organizations

limingcv's activity

upvoted a paper 6 days ago

Seedream 3.0 Technical Report

Paper • 2504.11346 • Published 7 days ago • 41

upvoted a paper 7 days ago

Have we unified image generation and understanding yet? An empirical study of GPT-4o's image generation ability

Paper • 2504.08003 • Published 13 days ago • 47

upvoted a paper 26 days ago

Wan: Open and Advanced Large-Scale Video Generative Models

Paper • 2503.20314 • Published 27 days ago • 49

upvoted 2 papers about 1 month ago

Unleashing Vecset Diffusion Model for Fast Shape Generation

Paper • 2503.16302 • Published Mar 20 • 44

Seedream 2.0: A Native Chinese-English Bilingual Image Generation Foundation Model

Paper • 2503.07703 • Published Mar 10 • 35

upvoted a paper 3 months ago

Cosmos World Foundation Model Platform for Physical AI

Paper • 2501.03575 • Published Jan 7 • 78

upvoted a paper 5 months ago

Enhancing the Reasoning Ability of Multimodal Large Language Models via Mixed Preference Optimization

Paper • 2411.10442 • Published Nov 15, 2024 • 81

upvoted 2 papers 6 months ago

GPT-4o System Card

Paper • 2410.21276 • Published Oct 25, 2024 • 86

Movie Gen: A Cast of Media Foundation Models

Paper • 2410.13720 • Published Oct 17, 2024 • 98

upvoted 2 papers 7 months ago

Molmo and PixMo: Open Weights and Open Data for State-of-the-Art Multimodal Models

Paper • 2409.17146 • Published Sep 25, 2024 • 114

OmniGen: Unified Image Generation

Paper • 2409.11340 • Published Sep 17, 2024 • 115

upvoted 2 papers 8 months ago

Imagen 3

Paper • 2408.07009 • Published Aug 13, 2024 • 62

CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer

Paper • 2408.06072 • Published Aug 12, 2024 • 40

upvoted 2 papers 9 months ago

Qwen2-Audio Technical Report

Paper • 2407.10759 • Published Jul 15, 2024 • 60

Qwen2 Technical Report

Paper • 2407.10671 • Published Jul 15, 2024 • 163

upvoted 2 papers 10 months ago

Depth Anything V2

Paper • 2406.09414 • Published Jun 13, 2024 • 103

An Image is Worth More Than 16x16 Patches: Exploring Transformers on Individual Pixels

Paper • 2406.09415 • Published Jun 13, 2024 • 52

upvoted 3 papers 11 months ago

Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation

Paper • 2406.06525 • Published Jun 10, 2024 • 71

Step-aware Preference Optimization: Aligning Preference with Denoising Performance at Each Step

Paper • 2406.04314 • Published Jun 6, 2024 • 30

ShareGPT4Video: Improving Video Understanding and Generation with Better Captions

Paper • 2406.04325 • Published Jun 6, 2024 • 76