[](#chatglm3-6b-32k)ChatGLM3-6B-32K
===================================

 [Github Repo](https://github.com/THUDM/ChatGLM)   [Twitter](https://twitter.com/thukeg)   [\[GLM@ACL 22\]](https://arxiv.org/abs/2103.10360) [\[GitHub\]](https://github.com/THUDM/GLM)   [\[GLM-130B@ICLR 23\]](https://arxiv.org/abs/2210.02414) [\[GitHub\]](https://github.com/THUDM/GLM-130B)  

 Join our [Slack](https://join.slack.com/t/chatglm/shared_invite/zt-25ti5uohv-A_hs~am_D3Q8XPZMpj7wwQ) and [WeChat](https://github.com/THUDM/ChatGLM/blob/main/resources/WECHAT.md)

Experience the larger-scale ChatGLM model at [chatglm.cn](https://www.chatglm.cn)

[](#-introduction) (Introduction)
-------------------------------------

ChatGLM3-6B-32K[ChatGLM3-6B](https://huggingface.co/THUDM/chatglm3-6b)32K 32K  **8K **[ChatGLM3-6B](https://huggingface.co/THUDM/chatglm3-6b)** 8K** ChatGLM3-6B-32K

ChatGLM3-6B  ChatGLM ChatGLM3-6B 

1.  **** ChatGLM3-6B  ChatGLM3-6B-Base ChatGLM3-6B-Base  10B 
2.  **** ChatGLM3-6B  [Prompt ](https://github.com/THUDM/ChatGLM3/blob/main/README.md)[](https://github.com/THUDM/ChatGLM3/blob/main/tools_using_demo/README.md)Function CallCode Interpreter Agent 
3.  ****  ChatGLM3-6B  ChatGLM-6B-Base ChatGLM3-6B-32K****[](https://open.bigmodel.cn/mla/form)****

Based on [ChatGLM3-6B](https://huggingface.co/THUDM/chatglm3-6b), ChatGLM3-6B-32K further strengthens the ability to understand long texts and can better handle contexts up to 32K in length. Specifically, we update the position encoding and design a more targeted long text training method, using a context length of 32K for training in the conversation stage. In actual use, if the context length you face is basically within **8K**, we recommend using [ChatGLM3-6B](https://huggingface.co/THUDM/chatglm3-6b); if you need to handle **For context lengths exceeding 8K**, we recommend using ChatGLM3-6B-32K.

ChatGLM3-6B is the latest open-source model in the ChatGLM series. While retaining many excellent features such as smooth dialogue and low deployment threshold from the previous two generations, ChatGLM3-6B introduces the following features:

1.  **More Powerful Base Model:** The base model of ChatGLM3-6B, ChatGLM3-6B-Base, employs a more diverse training dataset, more sufficient training steps, and a more reasonable training strategy. Evaluations on datasets such as semantics, mathematics, reasoning, code, knowledge, etc., show that ChatGLM3-6B-Base has the strongest performance among pre-trained models under 10B.
2.  **More Comprehensive Function Support:** ChatGLM3-6B adopts a newly designed [Prompt format](https://github.com/THUDM/ChatGLM3/blob/main/PROMPT_en.md), in addition to the normal multi-turn dialogue. It also natively supports [function call](https://github.com/THUDM/ChatGLM3/blob/main/tools_using_demo/README.md), code interpreter, and complex scenarios such as agent tasks.
3.  **More Comprehensive Open-source Series:** In addition to the dialogue model ChatGLM3-6B, the base model ChatGLM-6B-Base and the long-text dialogue model ChatGLM3-6B-32K are also open-sourced. All the weights are **fully open** for academic research, and after completing the [questionnaire](https://open.bigmodel.cn/mla/form) registration, they are also **allowed for free commercial use**.

[](#-dependencies) (Dependencies)
-----------------------------------------

    pip install protobuf transformers==4.30.2 cpm_kernels torch>=2.0 gradio mdtex2html sentencepiece accelerate
    

[](#-code-usage) (Code Usage)
-------------------------------------

 ChatGLM3-6B 

    >>> from transformers import AutoTokenizer, AutoModel
    >>> tokenizer = AutoTokenizer.from_pretrained("THUDM/chatglm3-6b-32k", trust_remote_code=True)
    >>> model = AutoModel.from_pretrained("THUDM/chatglm3-6b-32k", trust_remote_code=True).half().cuda()
    >>> model = model.eval()
    >>> response, history = model.chat(tokenizer, "", history=[])
    >>> print(response)
    ! ChatGLM-6B,,
    >>> response, history = model.chat(tokenizer, "", history=history)
    >>> print(response)
    ,:
    
    1. :,,
    2. :,,,
    3. :,,,,,
    4. :,,,
    5. :,,,
    6. :,,,,
    
    ,,
    

 DEMO [Github Repo](https://github.com/THUDM/ChatGLM)

For more instructions, including how to run CLI and web demos, and model quantization, please refer to our [Github Repo](https://github.com/THUDM/ChatGLM).

[](#-license) (License)
---------------------------

 [Apache-2.0](/THUDM/chatglm3-6b-32k/blob/main/LICENSE) ChatGLM3-6B  [Model License](/THUDM/chatglm3-6b-32k/blob/main/MODEL_LICENSE)

The code in this repository is open-sourced under the [Apache-2.0 license](/THUDM/chatglm3-6b-32k/blob/main/LICENSE), while the use of the ChatGLM3-6B model weights needs to comply with the [Model License](/THUDM/chatglm3-6b-32k/blob/main/MODEL_LICENSE).

[](#-citation) (Citation)
-----------------------------



If you find our work helpful, please consider citing the following papers.

    @article{zeng2022glm,
      title={Glm-130b: An open bilingual pre-trained model},
      author={Zeng, Aohan and Liu, Xiao and Du, Zhengxiao and Wang, Zihan and Lai, Hanyu and Ding, Ming and Yang, Zhuoyi and Xu, Yifan and Zheng, Wendi and Xia, Xiao and others},
      journal={arXiv preprint arXiv:2210.02414},
      year={2022}
    }
    

    @inproceedings{du2022glm,
      title={GLM: General Language Model Pretraining with Autoregressive Blank Infilling},
      author={Du, Zhengxiao and Qian, Yujie and Liu, Xiao and Ding, Ming and Qiu, Jiezhong and Yang, Zhilin and Tang, Jie},
      booktitle={Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers)},
      pages={320--335},
      year={2022}
    }

## Model Overview

The `chatglm3-6b-32k` is a large language model developed by THUDM. It is the latest open-source model in the ChatGLM series, which retains many excellent features from previous generations such as smooth dialogue and low deployment threshold, while introducing several key improvements.

Compared to the earlier [ChatGLM3-6B](https://aimodels.fyi/models/huggingFace/chatglm3-6b-thudm) model, `chatglm3-6b-32k` further strengthens the ability to understand long texts and can better handle contexts up to 32K in length. Specifically, the model updates the position encoding and uses a more targeted long text training method, with a context length of 32K during the conversation stage. This allows `chatglm3-6b-32k` to effectively process longer inputs compared to the 8K context length of `ChatGLM3-6B`.

The base model for `chatglm3-6b-32k`, called `ChatGLM3-6B-Base`, employs a more diverse training dataset, more training steps, and a refined training strategy. Evaluations show that `ChatGLM3-6B-Base` has the strongest performance among pre-trained models under 10B parameters on datasets covering semantics, mathematics, reasoning, code, and knowledge.

## Model Inputs and Outputs

### Inputs
- **Text**: The model can take text inputs of varying length, up to 32K tokens, and process them in a multi-turn dialogue setting.

### Outputs
- **Text response**: The model will generate relevant text responses based on the provided input and dialog history.

## Capabilities

`chatglm3-6b-32k` is a powerful language model that can engage in open-ended dialog, answer questions, provide explanations, and assist with a variety of language-based tasks. Some key capabilities include:

- **Long-form text understanding**: The model's 32K context length allows it to effectively process and reason about long-form inputs, making it well-suited for tasks involving lengthy documents or multi-turn conversations.
- **Multi-modal understanding**: In addition to regular text-based dialog, `chatglm3-6b-32k` also supports prompts that include functions, code, and other specialized inputs, allowing for more comprehensive task completion.
- **Strong general knowledge**: Evaluations show the underlying `ChatGLM3-6B-Base` model has impressive performance on a wide range of benchmarks, demonstrating broad and deep language understanding capabilities.

## What Can I Use It For?

The `chatglm3-6b-32k` model can be useful for a wide range of applications that require natural language processing and generation, especially those involving long-form text or multi-modal inputs. Some potential use cases include:

- **Conversational AI assistants**: The model's ability to engage in smooth, context-aware dialog makes it well-suited for building virtual assistants that can handle open-ended queries and maintain coherent conversations.
- **Content generation**: `chatglm3-6b-32k` can be used to generate high-quality text content, such as articles, reports, or creative writing, by providing appropriate prompts.
- **Question answering and knowledge exploration**: Leveraging the model's strong knowledge base, it can be used to answer questions, provide explanations, and assist with research and information discovery tasks.
- **Code generation and programming assistance**: The model's support for code-related inputs allows it to generate, explain, and debug code, making it a valuable tool for software development workflows.

## Things to Try

Some interesting things to try with `chatglm3-6b-32k` include:

- Engage the model in long-form, multi-turn conversations to test its ability to maintain context and coherence over extended interactions.
- Provide prompts that combine text with other modalities, such as functions or code snippets, to see how the model handles these more complex inputs.
- Explore the model's reasoning and problem-solving capabilities by giving it tasks that require analytical thinking, such as math problems or logical reasoning exercises.
- Fine-tune the model on domain-specific datasets to see how it can be adapted for specialized applications, like medical diagnosis, legal analysis, or scientific research.

By experimenting with the diverse capabilities of `chatglm3-6b-32k`, you can uncover new and innovative ways to leverage this powerful language model in your own projects and applications.