I benchmarked Aleph Alpha's Kolibri against Claude Sonnet 5.5 on long German writing: 0 of 136 blind wins, but $0.0077 an output on 2 rented GPUs
Kolibri (Aleph Alpha's open-weight German and English model: 78B total, 3.46B active, Apache 2.0) came out on 3 October. My app writes long German deep dives (about 1,200 to 1,500 words each) from source documents I put in the prompt. So I tested whether Kolibri could take over before switching any…
Read the full story at r/LocalLLM ↗