Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
3741.0
TFLOPS
Alexander Kozhevnikov
PRO
bethrezen
12
9
518
Follow
PhysiQuanty's profile picture
domofon's profile picture
21world's profile picture
17 followers
·
159 following
https://zeroagency.digital
bethrezen
AI & ML interests
LLM, TTS, SST, Recommenders
Recent Activity
liked
a dataset
about 23 hours ago
TeichAI/Ox-Alpha-10k
liked
a dataset
2 days ago
Roman1111111/GPT-5.6-luna-reasoning-102881x
reacted
to
FredyRivera-dev
's
post
with 👍
3 days ago
I've written a technical blog post about how we create a multimodal model: Kairos: a multimodal model built with LFM2.5-2.6B as the LLM, MoonViT-3D (the vision tower of Kimi-K2.6) as the vision encoder, and a custom projector. The original plan was LLaVA's approach, two stages: first align the projector with the LLM frozen, and then train the projector + LLM together. The first stage worked in terms of loss (ablation with +3.7 nats in favor of the image), but in free generation the image shifted the logits without changing the argmax: the model received the image and ignored it. That's why we jumped directly to early fusion, with a reasoning dataset. For that, we created Kairos-Multimodal-Reasoning: 116,357 examples with explicit reasoning traces, generated through distillation (60,041 from LLaVA-CC3M, 2,295 from WebSight, and 54,021 from Zebra-CoT), with GPT 5.6 Luna, Inkling, Qwen 3.6 27B, and Qwen 3.7 Plus as teachers. The training, in two phases: 1. Projector through backbone with 80k image-text pairs (Kairos-Proj-80k). 2. Projector + LoRA (r=16) with 30k examples from the reasoning dataset (Kairos-Alig-30k). Everything is open source: - Full blog post with the process: https://aquiles-ai.vercel.app/blog/kairos-a-multimodal-model - Implementation: https://github.com/Aquiles-ai/Kairos To be honest: the checkpoints are not a competent model, they are experimental artifacts. But they validated the approach and precisely defined what the next iteration needs. https://huggingface.co/collections/Aquiles-ai/kairos https://huggingface.co/Aquiles-ai/MoonViT-3D https://huggingface.co/LiquidAI/LFM2.5-2.6B
View all activity
Organizations
bethrezen
's activity
All
Models
Datasets
Spaces
Buckets
Papers
Collections
Community
Posts
Upvotes
Likes
Articles
liked
a dataset
about 23 hours ago
TeichAI/Ox-Alpha-10k
Viewer
•
Updated
about 13 hours ago
•
9.95k
•
6
liked
a dataset
2 days ago
Roman1111111/GPT-5.6-luna-reasoning-102881x
Preview
•
Updated
2 days ago
•
44
•
2
liked
3 datasets
3 days ago
MidTool/MidTool-Mix
Viewer
•
Updated
5 days ago
•
11.2M
•
19
•
2
ankushthakurr09/MimanusM1_v1_Dataset
Viewer
•
Updated
5 days ago
•
5.83M
•
87
•
2
codelion/sutra-10B
Viewer
•
Updated
Mar 8
•
5M
•
395
•
12
liked
a dataset
4 days ago
OpenResearcher/OpenResearcher-Corpus
Viewer
•
Updated
4 days ago
•
14.9M
•
1.75k
•
11
liked
3 datasets
6 days ago
Dahoas/full-hh-rlhf
Viewer
•
Updated
Feb 23, 2023
•
125k
•
964
•
91
renhuimin/RL-Instruction-Following-Dataset
Viewer
•
Updated
Jan 14
•
168k
•
141
•
4
PleIAs/SYNTH
Viewer
•
Updated
May 6
•
68M
•
17.9k
•
275
liked
a dataset
7 days ago
DatasetsEval/RusLang-edu-1000
Viewer
•
Updated
3 days ago
•
1k
•
80
•
1
liked
2 datasets
11 days ago
mlabonne/orca-agentinstruct-1M-v1-cleaned
Viewer
•
Updated
Jan 25, 2025
•
1.05M
•
293
•
69
Huang2020/qwen3.6-27B-reasoning-regen
Viewer
•
Updated
Jul 10
•
2.65M
•
259
•
3
liked
a model
11 days ago
Qwen/Qwen3.8-27B
Image-Text-to-Text
•
28B
•
Updated
11 days ago
•
2.65M
•
•
12.6k
liked
a dataset
11 days ago
ChengyuDu0123/HER-Dataset
Viewer
•
Updated
Feb 4
•
419k
•
557
•
14
liked
2 datasets
13 days ago
Bas95/reasoning-distill-claude-opus-4-7-max
Viewer
•
Updated
Apr 29
•
8.12k
•
128
•
3
nvidia/Nemotron-SFT-Math-v4
Viewer
•
Updated
13 days ago
•
545k
•
5.28k
•
38
liked
2 datasets
19 days ago
IIGroup/X-Coder-SFT-376k
Viewer
•
Updated
Feb 7
•
887k
•
1.47k
•
21
mizinovmv/GLM-5.2-Finance-80000x-ru
Viewer
•
Updated
19 days ago
•
78.7k
•
86
•
1
liked
a model
19 days ago
deepgrove/maple-preview
Text Generation
•
20B
•
Updated
20 days ago
•
7.18k
•
376
liked
a dataset
20 days ago
hotpotqa/hotpot_qa
Viewer
•
Updated
Aug 11, 2025
•
203k
•
84.9k
•
325
Load more