Base Model: replit/replit-code-v1-3b

This is version 2 of the Replit Code Instruct fine tune model.

This model is fine tuned on both Sahil2801's CodeAlpaca & Teknium's GPTeacher Code-Instruct to give Replit's Code model instruct capabilities.

Try this model on it's HuggingFace demo Spaces: [https://huggingface.co/spaces/teknium/Replit-v2-CodeInstruct-3B](https://huggingface.co/spaces/teknium/Replit-v2-CodeInstruct-3B)

Dataset links: CodeAlpaca: [https://huggingface.co/datasets/sahil2801/CodeAlpaca-20k](https://huggingface.co/datasets/sahil2801/CodeAlpaca-20k) GPTeacher subset - Code Instruct: [https://github.com/teknium1/GPTeacher](https://github.com/teknium1/GPTeacher)

This model was trained on 2x a100 80gb for 1 hour on ~25,000 code instruction/response pairs in Alpaca format.

The first model was only trained on 512 sequence length, this model on 2000, giving it much greater access to training data knowledge.

Refer to the base models HuggingFace model card for some basic requirements to run: [https://huggingface.co/replit/replit-code-v1-3b](https://huggingface.co/replit/replit-code-v1-3b)

This fine tune can be prompted like any alpaca fine tune:

    ### Instruction:
    <prompt>
    
    ### Input:
    <additional context>
    
    ### Response:
    

or

    ### Instruction:
    <prompt>
    
    ### Response:
    

This model seems to have issues with device="auto" in the model arguments (and requires the trust\_remote\_code=True, so you should maybe load it like I am here:

            self.tokenizer = AutoTokenizer.from_pretrained("./Replit-CodeInstruct/", trust_remote_code=True)
            self.model = AutoModelForCausalLM.from_pretrained(
                "./Replit-CodeInstruct",
                torch_dtype=torch.bfloat16,
                trust_remote_code=True
            )
            self.model.to('cuda')
    

This model for me produced coherent outputs with the following sampler settings, but feel free to experiment:

    max_new_tokens=128, do_sample=True, use_cache=True, temperature=0.2, top_p=0.9, eos_token_id= self.tokenizer.eos_token_id
    

In the tokenizer decode arguments, it also needs these settings:

    skip_special_tokens=True, clean_up_tokenization_space=False
    

The following parameters were used with HuggingFace trainer to train the model with:

    --model_name_or_path replit/replit-code-v1-3b --data_path /root/stanford_alpaca/train.json --bf16 True --output_dir /root/stanford_alpaca/model_ckpts --num_train_epochs 3 --per_device_train_batch_size 4 --per_device_eval_batch_size 1 --gradient_accumulation_steps 8 --save_strategy steps --save_steps 200 --save_total_limit 3 --learning_rate 1e-5 --weight_decay 0. --warmup_ratio 0.03 --tf32 True --run_name Replit1

## Model Overview

The `Replit-v2-CodeInstruct-3B` model is a 3 billion parameter AI model developed by [teknium](https://aimodels.fyi/creators/huggingFace/teknium) that has been fine-tuned on both the [CodeAlpaca](https://huggingface.co/datasets/sahil2801/CodeAlpaca-20k) and [GPTeacher Code-Instruct](https://github.com/teknium1/GPTeacher) datasets to give it code instruction capabilities. This model builds upon the [replit-code-v1-3b](https://aimodels.fyi/models/huggingFace/replit-code-v1-3b-replit) base model, which was trained on a diverse set of programming languages. The fine-tuning process has given the `Replit-v2-CodeInstruct-3B` model the ability to follow code-related instructions and generate relevant responses.

## Model Inputs and Outputs

### Inputs
- **Code-related prompts and instructions**: The model is designed to accept text-based prompts and instructions related to coding tasks, such as "Write a function that computes the Fibonacci sequence up to n" or "Explain how this code snippet works."

### Outputs
- **Generated code and text responses**: The model can generate relevant code snippets and text-based responses to address the provided instructions and prompts. The outputs aim to be helpful, informative, and aligned with the user's intent.

## Capabilities

The `Replit-v2-CodeInstruct-3B` model is capable of engaging in a wide range of code-related tasks, such as code completion, code explanation, and generating code based on natural language instructions. It can handle prompts across multiple programming languages, including Python, JavaScript, Java, and more. The model's fine-tuning on the CodeAlpaca and GPTeacher datasets has improved its ability to follow instructions and provide helpful, coherent responses.

## What Can I Use It For?

The `Replit-v2-CodeInstruct-3B` model can be a valuable tool for developers and researchers working on projects that involve code generation, code understanding, and code-related task completion. It can be used to build applications that assist programmers by providing code suggestions, explanations, and solutions to coding problems. Additionally, the model could be further fine-tuned or integrated into educational resources or coding learning tools to support students and beginners in their programming journeys.

## Things to Try

One interesting thing to try with the `Replit-v2-CodeInstruct-3B` model is to explore its ability to handle code-related prompts that involve multiple steps or complex instructions. For example, you could try asking the model to write a function that solves a specific coding challenge, or to explain the inner workings of a given code snippet in detail. Experimenting with different types of prompts and observing the model's responses can help you better understand its capabilities and limitations.