Skip to content

WizardCoder: Empowering Code Large Language Models with Evol-Instruct research paper by Microsoft, 2023

Microsoft · Jun 14, 2023 · Foundation models · 34 upvotes · unverified 3 years ago

Read on arXiv

What it shows

WizardCoder, a Code LLM fine-tuned with complex instructions using Evol-Instruct, outperforms other open-source and closed LLMs on several code generation benchmarks.

UnverifiedHugging Face's summary; not yet checked by hand.

More from Microsoft

All 47
PaperCitations
Agensh: Scaling Organizational Intelligence to 1,024 AgentsA multi-agent system can reduce latency on complex tasks by executing work concurrently.Agents and evaluation · Sep 2026 · Unverified6 days ago-
The Tasteful Agent: Measuring and Improving Taste in Long-Horizon TasksLLM agents increasingly work on long-horizon tasks, and the decisions they make along the way, such as which hypothesis to test or which implementation to build on, determine the outcome of the whole run.Agents and evaluation · Sep 2026 · Unverified6 days ago0
When EOS Tokens Disagree: Understanding Length Inflation in On-Policy DistillationWe study length inflation in on-policy distillation (OPD), where student responses can become excessively long and even exhaust the generation budget.Foundation models · Sep 2026 · Unverified11 days ago0
When2Think: Learning Difficulty-Aware Length Control for Efficient Hybrid Reasoning ModelsLarge Reasoning Models (LRMs) achieve strong performance on complex tasks but exhibit systematic inefficiency: they often overthink easy problems and underthink hard ones.Reasoning · Sep 2026 · Unverified11 days ago0
BI-Agent and BI-Bench: Towards Automating End-to-End Business IntelligenceBusiness intelligence (BI) is a cornerstone of enterprise decision-making and is widely used by enterprise users in software such as Power BI and Tableau.Agents and evaluation · Sep 2026 · Unverified12 days ago0
StudentSim: Training LLM-based Student SimulatorsStudentSim trains per-student simulators that answer like a given learner and change their answers under a tutor's guidance.Applied AI · Sep 20263 weeks ago0
AutoSaddler: Automatic Harness Optimization with Durable Updates from Agent Execution TracesAutoSaddler automatically improves LLM agent harnesses via offline failure-driven optimization, boosting performance on long-horizon benchmarks.Agents and evaluation · Aug 2026 · Unverified5 weeks ago4
Agent Lightning v1.0: Towards Harnessed Agentic RLAgent Lightning v1.0 enables reproducible reinforcement learning for arbitrary agent harnesses, substantially improving coding-agent performance with minimal data and compute.Agents and evaluation · Aug 2026 · Unverified5 weeks ago2
Topic
About this paper
Authors
Ziyang Luo, Can Xu, Pu Zhao and 7 more
arXiv
2306.08568 · PDF
Citations
Not counted yet · Semantic Scholar
Upvotes
34 · Hugging Face
Code
github.com/nlpxucan/WizardLM
Lab
Microsoft · on Companies · on Acquisitions · on Quarterly · on Paydays · on Releases · on TechConf

Changes

What changed
New paperFound by the weekly scan, unverifiedSep 28, 2026today

New papers by email

Monday afternoons, only in weeks with new papers from the labs.

Double opt-in. Unsubscribe any time.