search
Get Started
search
Yi-68B - Self Hosted
zoom_in Click to enlarge

Yi-68B

language

description Yi-68B Overview

Yi-68B is a powerful open-source LLM developed by 01.AI. It stands out for its exceptionally long context window (up to 32K tokens), enabling it to process and generate text based on significantly larger inputs than many competitors. This makes it ideal for tasks requiring deep understanding of complex documents or conversations.

help Yi-68B FAQ

Who created the Yi-68B language model?

Yi-68B was released by 01.AI, the company founded by AI researcher Kai-Fu Lee. It belongs to the original Yi family alongside the smaller Yi-6B model.

How much hardware is needed to run Yi-68B locally?

Running all 68 billion parameters at 16-bit precision requires well over 130 GB of memory before runtime overhead. A 4-bit quantization is much smaller but still typically needs roughly 40 GB or more across GPU memory and system RAM.

What context length does Yi-68B support?

The original Yi-68B release supports a context window of up to 4K tokens, while the separately released Yi-68B-200K extends that to 200K. A catalog claim of a universal 32K limit therefore does not identify the model variants correctly.

Is Yi-68B a chat model or a base model?

Both forms were released: Yi-68B is the pretrained base model, and Yi-68B-Chat is tuned for conversational instruction following. Self-hosters should choose the chat variant for an assistant interface and the base model for further training or completion workflows.

Reviews & Comments

Write a Review

rate_review

Be the first to review

Share your thoughts with the community and help others make better decisions.

Save to your list

Save your favorites and follow how their scores change over time.

Save favorites
Track changes
Compare scores

Already have an account? Sign in

Compare Items

See how they stack up against each other

Comparing
VS
Select 1 more item to compare