Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up

tensorblock
/
Einstein-v7-Qwen2-7B-GGUF

GGUF
English
axolotl
instruct
finetune
chatml
gpt4
synthetic data
science
physics
chemistry
biology
math
qwen
qwen2
TensorBlock
GGUF
Eval Results (legacy)
conversational
Model card Files Files and versions
xet
Community

Instructions to use tensorblock/Einstein-v7-Qwen2-7B-GGUF with libraries, inference providers, notebooks, and local apps. Follow these links to get started.

  • Notebooks
  • Google Colab
  • Kaggle
  • Local Apps Settings
  • llama.cpp

    How to use tensorblock/Einstein-v7-Qwen2-7B-GGUF with llama.cpp:

    Install (macOS, Linux)
    curl -LsSf https://llama.app/install.sh | sh
    # Start a local OpenAI-compatible server with a web UI:
    llama serve -hf tensorblock/Einstein-v7-Qwen2-7B-GGUF:Q2_K
    # Run inference directly in the terminal:
    llama cli -hf tensorblock/Einstein-v7-Qwen2-7B-GGUF:Q2_K
    Install from WinGet (Windows)
    winget install llama.cpp
    # Start a local OpenAI-compatible server with a web UI:
    llama serve -hf tensorblock/Einstein-v7-Qwen2-7B-GGUF:Q2_K
    # Run inference directly in the terminal:
    llama cli -hf tensorblock/Einstein-v7-Qwen2-7B-GGUF:Q2_K
    Use pre-built binary
    # Download pre-built binary from:
    # https://github.com/ggerganov/llama.cpp/releases
    # Start a local OpenAI-compatible server with a web UI:
    ./llama-server -hf tensorblock/Einstein-v7-Qwen2-7B-GGUF:Q2_K
    # Run inference directly in the terminal:
    ./llama-cli -hf tensorblock/Einstein-v7-Qwen2-7B-GGUF:Q2_K
    Build from source code
    git clone https://github.com/ggerganov/llama.cpp.git
    cd llama.cpp
    cmake -B build
    cmake --build build -j --target llama-server llama-cli
    # Start a local OpenAI-compatible server with a web UI:
    ./build/bin/llama-server -hf tensorblock/Einstein-v7-Qwen2-7B-GGUF:Q2_K
    # Run inference directly in the terminal:
    ./build/bin/llama-cli -hf tensorblock/Einstein-v7-Qwen2-7B-GGUF:Q2_K
    Use Docker
    docker model run hf.co/tensorblock/Einstein-v7-Qwen2-7B-GGUF:Q2_K
  • LM Studio
  • Jan
  • Ollama

    How to use tensorblock/Einstein-v7-Qwen2-7B-GGUF with Ollama:

    ollama run hf.co/tensorblock/Einstein-v7-Qwen2-7B-GGUF:Q2_K
  • Unsloth Desktop
  • Docker Model Runner

    How to use tensorblock/Einstein-v7-Qwen2-7B-GGUF with Docker Model Runner:

    docker model run hf.co/tensorblock/Einstein-v7-Qwen2-7B-GGUF:Q2_K
  • Lemonade

    How to use tensorblock/Einstein-v7-Qwen2-7B-GGUF with Lemonade:

    Pull the model
    # Download Lemonade from https://lemonade-server.ai/
    lemonade pull tensorblock/Einstein-v7-Qwen2-7B-GGUF:Q2_K
    Run and chat with the model
    lemonade run user.Einstein-v7-Qwen2-7B-GGUF-Q2_K
    List all available models
    lemonade list
  • Atomic Chat
Einstein-v7-Qwen2-7B-GGUF
6.82 GB
Ctrl+K
Ctrl+K
  • 1 contributor
History: 7 commits
morriszms's picture
morriszms
Keep Q2_K/Q3_K_M gguf only
56b8511 verified 7 months ago
  • .gitattributes
    2.34 kB
    Upload folder using huggingface_hub almost 2 years ago
  • Einstein-v7-Qwen2-7B-Q2_K.gguf
    3.02 GB
    xet
    Upload folder using huggingface_hub almost 2 years ago
  • Einstein-v7-Qwen2-7B-Q3_K_M.gguf
    3.81 GB
    xet
    Upload folder using huggingface_hub almost 2 years ago
  • README.md
    10.4 kB
    Update README.md about 1 year ago