Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up

Aleph-Alpha
/
Kolibri-1

Text Generation
Safetensors
German
English
vllm
kolibri1
reasoning
Mixture of Experts
conversational
Eval Results
fp8
Model card Files Files and versions
xet
Community
11
New discussion
Resources
  • PR & discussions documentation
  • Code of Conduct
  • Hub documentation

Add bibtex entry for citation

#11 opened 1 day ago by
jonasknupp

[Evaluation Report] Kolibri-1 FP8 on one RTX PRO 6000 (96 GB): serving numbers, and BEAM 100K from its own 262K window vs a 7K-token memory page

#10 opened 2 days ago by
Robin1986

Tool Call caching in ram breaks the model

#9 opened 3 days ago by
Dreadbyte

Broken tool-call parser in the ghcr.io/aleph-alpha/aleph-alpha-inference image

3
#8 opened 4 days ago by
triviumzh

[Evaluation Report] Empirical NIS-2 Stress Test: Closed-Book Drift vs. RAG Grounding on H200 (1,850 Runs)

1
#7 opened 4 days ago by
sbeierle

Add evaluation results

#6 opened 5 days ago by
SaylorTwift

DGX Spark recipe

๐Ÿ‘ 2
#5 opened 7 days ago by
stelterlab

tool-eval-bench results

๐Ÿ‘ 6
1
#4 opened 7 days ago by
stelterlab

llama.cpp support?

โž• 6
4
#3 opened 7 days ago by
jacek2024

transformers support for Kolibri1ForCausalLM

2
#2 opened 7 days ago by
stelterlab

Request: Quantized versions (MLX / GGUF) for Apple Silicon โ€“ happy to benchmark

๐Ÿ‘ 10
6
#1 opened 7 days ago by
Blackbeard82
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs