ChatGLM-6B

ChatGLM-6B is an open bilingual (Chinese + English) conversational language model based on the GLM architecture, with approximately 6.2 billion parameters. The project provides inference code, demos (command line, web, API), quantization support for lower memory deployment, and tools for finetuning (e.g., via P-Tuning v2). It is optimized for dialogue and question answering with a balance between performance and deployability in consumer hardware settings. Support for quantized inference (INT4, INT8) to reduce GPU memory requirements. Automatic mode switching between precision/memory tradeoffs (full/quantized).

Features

Bilingual dialogue capability (Chinese and English)
Support for quantized inference (INT4, INT8) to reduce GPU memory requirements
Parameter-efficient finetuning (P-Tuning v2 method)
CLI, web demo, and API interfaces included
Automatic mode switching between precision / memory tradeoffs (full / quantized)
Integration with the Hugging Face / Transformers ecosystem (trust_remote_code, model loading, tokenization)

Project Samples

Project Activity

See All Activity >

License

Apache License V2.0

Follow ChatGLM-6B

ChatGLM-6B Web Site

Other Useful Business Software

Keep company data safe with Chrome Enterprise

Protect your business with AI policies and data loss prevention in the browser

Make AI work your way with Chrome Enterprise. Block unapproved sites and set custom data controls that align with your company's policies.

Download Chrome

Rate This Project

User Reviews

Be the first to post a review of ChatGLM-6B!

Additional Project Details

Programming Language

Python

Related Categories

Python Large Language Models (LLM), Python AI Models

Registered

2025-09-26

Similar Business Software

ChatGLM

ChatGLM-6B is an open-source, Chinese-English bilingual dialogue language model based on the General Language Model (GLM) architecture with 6.2 billion parameters. Combined with model quantization technology, users can deploy locally on consumer-grade graphics cards (only 6GB of video memory is...

See Software
GLM-4.5

GLM‑4.5 is Z.ai’s latest flagship model in the GLM family, engineered with 355 billion total parameters (32 billion active) and a companion GLM‑4.5‑Air variant (106 billion total, 12 billion active) to unify advanced reasoning, coding, and agentic capabilities in one architecture. It operates in...

See Software
Baichuan-13B

Baichuan-13B is an open source and commercially available large-scale language model containing 13 billion parameters developed by Baichuan Intelligent following Baichuan -7B . It has achieved the best results of the same size on authoritative Chinese and English benchmarks. This release...

See Software
DeepSeek-V3.1-Terminus

DeepSeek has released DeepSeek-V3.1-Terminus, which enhances the V3.1 architecture by incorporating user feedback to improve output stability, consistency, and agent performance. It notably reduces instances of mixed Chinese/English character output and unintended garbled characters, resulting...

See Software
Qwen

Qwen LLM refers to a family of large language models (LLMs) developed by Alibaba Cloud's Damo Academy. These models are trained on a massive dataset of text and code, allowing them to understand and generate human-like text, translate languages, write different kinds of creative content, and...

See Software
DeepSeek-V2

DeepSeek-V2 is a state-of-the-art Mixture-of-Experts (MoE) language model introduced by DeepSeek-AI, characterized by its economical training and efficient inference capabilities. With a total of 236 billion parameters, of which only 21 billion are active per token, it supports a context length...

See Software

Report inappropriate content

ChatGLM-6B

ChatGLM-6B: An Open Bilingual Dialogue Language Model

Get an email when there's a new version of ChatGLM-6B

Features

Project Samples

Project Activity

Categories

License

Follow ChatGLM-6B

User Reviews

Additional Project Details

Programming Language

Related Categories

Registered