Athrael Soju PRO
athrael-soju
AI & ML interests
Yes
Recent Activity
new activity 5 days ago
vultr/VultronRetrieverFlash-Qwen3.5-0.8B:Embeddings differ between colpali engine & vLLM updated a model 5 days ago
athrael-soju/colqwen3.5-4.5B-v3 updated a model 5 days ago
vultr/VultronRetrieverPrime-Qwen3.5-8BOrganizations
Hydra — Dual-Head Retrieval and Generation
Dual-head VLM: ColBERT retrieval + autoregressive generation by toggling one LoRA. Canonical 4B + 0.8B, omni proof-of-concept, baselines.
ColGemma4 — Gemma-4 Visual Retrieval
ColBERT-style late-interaction visual document retrieval adapters built on Google Gemma-4 (E2B and E4B variants).
Models
-
athrael-soju/colqwen3.5-4.5B-v3
Visual Document Retrieval • 5B • Updated • 99.1k • 12 -
athrael-soju/HydraQwen3.5-0.8B
Visual Document Retrieval • Updated • 14 -
athrael-soju/ColGemma4-E4B-IT-Base
Visual Document Retrieval • Updated • 1 -
athrael-soju/ColGemma4-E2B-IT-Base
Visual Document Retrieval • Updated • 1
ColQwen3.5 — Qwen3.5 Visual Retrieval
Visual document retrieval models on Qwen3.5 backbone. ViDoRe v3 leaderboard competitors, 128-dim multi-vector.
Papers
-
The Price of Anarchy in Disaggregated Inference
Paper • 2606.17081 • Published • 1 -
Spatially-Grounded Document Retrieval via Patch-to-Region Relevance Propagation
Paper • 2512.02660 • Published -
Architecture-Aware LLM Inference Optimization on AMD Instinct GPUs: A Comprehensive Benchmark and Deployment Study
Paper • 2603.10031 • Published -
Hydra: Unifying Document Retrieval and Generation in a Single Vision-Language Model
Paper • 2603.28554 • Published
Vultr Models
Models
-
athrael-soju/colqwen3.5-4.5B-v3
Visual Document Retrieval • 5B • Updated • 99.1k • 12 -
athrael-soju/HydraQwen3.5-0.8B
Visual Document Retrieval • Updated • 14 -
athrael-soju/ColGemma4-E4B-IT-Base
Visual Document Retrieval • Updated • 1 -
athrael-soju/ColGemma4-E2B-IT-Base
Visual Document Retrieval • Updated • 1
Hydra — Dual-Head Retrieval and Generation
Dual-head VLM: ColBERT retrieval + autoregressive generation by toggling one LoRA. Canonical 4B + 0.8B, omni proof-of-concept, baselines.
ColQwen3.5 — Qwen3.5 Visual Retrieval
Visual document retrieval models on Qwen3.5 backbone. ViDoRe v3 leaderboard competitors, 128-dim multi-vector.
ColGemma4 — Gemma-4 Visual Retrieval
ColBERT-style late-interaction visual document retrieval adapters built on Google Gemma-4 (E2B and E4B variants).
Papers
-
The Price of Anarchy in Disaggregated Inference
Paper • 2606.17081 • Published • 1 -
Spatially-Grounded Document Retrieval via Patch-to-Region Relevance Propagation
Paper • 2512.02660 • Published -
Architecture-Aware LLM Inference Optimization on AMD Instinct GPUs: A Comprehensive Benchmark and Deployment Study
Paper • 2603.10031 • Published -
Hydra: Unifying Document Retrieval and Generation in a Single Vision-Language Model
Paper • 2603.28554 • Published