4 1

Soham Parikh

sohampnow

sohampnow

AI & ML interests

None yet

Recent Activity

new activity 21 days ago

allenai/OLMoE-1B-7B-0924:Going deeper vs wider

new activity 22 days ago

allenai/OLMoE-1B-7B-0924:Reason for norm_topk_prob=false?

View all activity

Organizations

sohampnow's activity

New activity in allenai/OLMoE-1B-7B-0924 21 days ago

Going deeper vs wider

#5 opened 21 days ago by

sohampnow

New activity in allenai/OLMoE-1B-7B-0924 22 days ago

Reason for norm_topk_prob=false?

#4 opened 23 days ago by

sohampnow

reacted to WizardLM's post with 🚀 5 months ago

Post

10514

🔥 🔥🔥
Excited to announce WizardLM new Paper: Auto Evol-Instruct!

🐦 Twitter: https://x.com/WizardLM_AI/status/1812857977122202087

📃 Paper: https://arxiv.org/pdf/2406.00770

🤖 1. Fully AI-Powered Pipeline

Auto Evol-Instruct automatically involves an iterative process of optimizing an Evol-Instruct V1 into an optimal one. The pipeline consists of two critical stages: Evol Trajectory Analysis, where the optimizer LLM analyzes the issues and failures exposed in instruction evolution performed by the evol LLM, and Evolving Method Optimization, where the optimizer LLM addresses these issues to progressively develop an effective evolving method. The optimal evolving method is then used to convert the entire instruction dataset into more diverse and complex forms, facilitating improved instruction tuning.

📈2. Scaling Evol-Instruct with Arena Learning

With Auto Evol-Instruct, the evolutionary synthesis data of WizardLM-2 has scaled up from WizardLM-1 to dozens of domains, covering tasks in all aspects of large language models. This allows Arena Learning to train and learn from an almost infinite pool of high-difficulty instruction data, fully unlocking all the potential of Arena Learning.