Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
3741.0
TFLOPS
Alexander Kozhevnikov
PRO
bethrezen
12
9
518
Follow
domofon's profile picture
Roman190928's profile picture
PhysiQuanty's profile picture
17 followers
·
159 following
https://zeroagency.digital
bethrezen
AI & ML interests
LLM, TTS, SST, Recommenders
Recent Activity
liked
a dataset
37 minutes ago
TeichAI/ox-alpha-10k
liked
a dataset
about 20 hours ago
Roman1111111/GPT-5.6-luna-reasoning-102881x
reacted
to
FredyRivera-dev
's
post
with 👍
2 days ago
I've written a technical blog post about how we create a multimodal model: Kairos: a multimodal model built with LFM2.5-2.6B as the LLM, MoonViT-3D (the vision tower of Kimi-K2.6) as the vision encoder, and a custom projector. The original plan was LLaVA's approach, two stages: first align the projector with the LLM frozen, and then train the projector + LLM together. The first stage worked in terms of loss (ablation with +3.7 nats in favor of the image), but in free generation the image shifted the logits without changing the argmax: the model received the image and ignored it. That's why we jumped directly to early fusion, with a reasoning dataset. For that, we created Kairos-Multimodal-Reasoning: 116,357 examples with explicit reasoning traces, generated through distillation (60,041 from LLaVA-CC3M, 2,295 from WebSight, and 54,021 from Zebra-CoT), with GPT 5.6 Luna, Inkling, Qwen 3.6 27B, and Qwen 3.7 Plus as teachers. The training, in two phases: 1. Projector through backbone with 80k image-text pairs (Kairos-Proj-80k). 2. Projector + LoRA (r=16) with 30k examples from the reasoning dataset (Kairos-Alig-30k). Everything is open source: - Full blog post with the process: https://aquiles-ai.vercel.app/blog/kairos-a-multimodal-model - Implementation: https://github.com/Aquiles-ai/Kairos To be honest: the checkpoints are not a competent model, they are experimental artifacts. But they validated the approach and precisely defined what the next iteration needs. https://huggingface.co/collections/Aquiles-ai/kairos https://huggingface.co/Aquiles-ai/MoonViT-3D https://huggingface.co/LiquidAI/LFM2.5-2.6B
View all activity
Organizations
bethrezen
's models
6
Sort: Recently updated
bethrezen/gpt-oss-20b-merged
Text Generation
•
21B
•
Updated
Aug 8, 2025
•
8
bethrezen/gpt-oss-20b-multilingual-reasoner
Updated
Aug 8, 2025
bethrezen/zero-summary-v2-beta15-e1-Q8_0-GGUF
8B
•
Updated
May 14, 2025
•
11
bethrezen/zero-summary-v2-beta15-Q4_K_M-GGUF
8B
•
Updated
May 14, 2025
•
9
bethrezen/zero-summary-v2-beta3-Q8_0-GGUF
24B
•
Updated
May 14, 2025
•
11
bethrezen/zero-summary-v2-beta15-Q8_0-GGUF
8B
•
Updated
May 11, 2025
•
10