Dl's picture

55 599

Dl

Dlbk

·

AI & ML interests

None yet

Recent Activity

liked a model 2 days ago

moonshotai/Kimi-VL-A3B-Instruct

liked a model 2 days ago

moonshotai/Kimi-VL-A3B-Thinking

upvoted a collection 2 days ago

View all activity

Organizations

Dlbk's activity

upvoted a collection 2 days ago

Kimi-VL-A3B

Moonshot's efficient MoE VLMs, exceptional on agent, long-context, and thinking • 6 items • Updated about 2 hours ago • 57

upvoted a collection 2 months ago

Moshi v0.1 Release

MLX, Candle & PyTorch model checkpoints released as part of the Moshi release from Kyutai. Run inference via: https://github.com/kyutai-labs/moshi • 13 items • Updated Sep 18, 2024 • 227

upvoted an article 2 months ago

Article

Open-source DeepResearch – Freeing our search agents

Feb 4

• 1.21k

upvoted a collection 2 months ago

Qwen2.5-VL

Vision-language model series based on Qwen2.5 • 11 items • Updated 12 days ago • 441

upvoted a paper 3 months ago

Demons in the Detail: On Implementing Load Balancing Loss for Training Specialized Mixture-of-Expert Models

Paper • 2501.11873 • Published Jan 21 • 66

upvoted 2 collections 3 months ago

OuteTTS

10 items • Updated 5 days ago • 15

OuteTTS 0.3

4 items • Updated 5 days ago • 18

upvoted a paper 3 months ago

MiniMax-01: Scaling Foundation Models with Lightning Attention

Paper • 2501.08313 • Published Jan 14 • 285

upvoted 3 collections 4 months ago

QwQ

Qwen with Questions • 6 items • Updated Mar 6 • 93

QVQ

QVQ: Qwen models for visual reasoning • 7 items • Updated Jan 1 • 49

Falcon3

Falcon3 family of Open Foundation Models is a set of pretrained and instruct LLMs ranging from 1B to 10B parameters. • 40 items • Updated Feb 13 • 84

upvoted a paper 4 months ago

Apollo: An Exploration of Video Understanding in Large Multimodal Models

Paper • 2412.10360 • Published Dec 13, 2024 • 147

upvoted 2 collections 4 months ago

GLM-4

GLM-4 Open Models • 14 items • Updated 4 days ago • 117

DeepSeek-V2.5

2 items • Updated Dec 10, 2024 • 40

upvoted a paper 4 months ago

o1-Coder: an o1 Replication for Coding

Paper • 2412.00154 • Published Nov 29, 2024 • 45

upvoted 2 collections 7 months ago

Qwen2.5-Coder

Code-specific model series based on Qwen2.5 • 40 items • Updated Nov 28, 2024 • 302

Qwen2.5

Qwen2.5 language models, including pretrained and instruction-tuned models of 7 sizes, including 0.5B, 1.5B, 3B, 7B, 14B, 32B, and 72B. • 46 items • Updated Feb 26 • 587

upvoted an article 7 months ago

Article

The Open Medical-LLM Leaderboard: Benchmarking Large Language Models in Healthcare

Apr 19, 2024

• 144

upvoted a collection 8 months ago

Qwen2-VL

Vision-language model series based on Qwen2 • 16 items • Updated Dec 6, 2024 • 211