Kaveri-4B

Kaveri-4B is a compact coding-focused language model fine-tuned for software engineering, algorithmic problem solving, competitive programming, and Python code generation.

Base Model

Kaveri-4B is derived from Qwen3.5-4B.

Fine-tuning

  • Method: 16-bit LoRA
  • LoRA rank: 16
  • LoRA alpha: 16
  • Effective batch size: 8
  • Training steps: 500
  • Training data: execution-filtered NVIDIA OpenCodeInstruct
  • Quality threshold: >= 0.80
  • Focus: code generation and algorithmic reasoning

Evaluation

The model is intended to be evaluated using LiveCodeBench, including Pass@1 code-generation evaluation.

Intended Uses

  • code generation
  • algorithm implementation
  • competitive programming
  • programming assistance
  • coding-model research

Model Name

Kaveri-4B

Downloads last month
282
Safetensors
Model size
5B params
Tensor type
BF16
ยท
F32
ยท
Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for dharun2049/Kaveri-4B

Finetuned
Qwen/Qwen3.5-4B
Adapter
(642)
this model
Adapters
1 model

Space using dharun2049/Kaveri-4B 1