Skip to content
View Orion-zhen's full-sized avatar
💥
CUDA Out Of Memory
💥
CUDA Out Of Memory

Block or report Orion-zhen

Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Pinned Loading

  1. our our Public

    Orion User's Repository for Arch Linux

    Svelte 2

  2. abliteration abliteration Public

    Make abliterated models with transformers, easy and fast

    Python 166 90

  3. turboderp-org/exllamav2 turboderp-org/exllamav2 Public

    A fast inference library for running LLMs locally on modern consumer-class GPUs

    Python 4.6k 341

  4. hiyouga/LlamaFactory hiyouga/LlamaFactory Public

    Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)

    Python 74.5k 9.1k

  5. Tiiny-AI/PowerInfer Tiiny-AI/PowerInfer Public

    High-speed Large Language Model Serving for Local Deployment

    C++ 9.8k 597

  6. CrazyBoyM/llama3-Chinese-chat CrazyBoyM/llama3-Chinese-chat Public

    Llama3-中文后训练版

    Python 4.1k 331