EveryDev.ai
Subscribe
Home
Developers

3,067+ AI companies

  • Radar
  • Trending
AI Tools by Topic
  • AI Coding Assistants
  • Agent Frameworks
  • MCP Servers
  • AI Prompt Tools
  • Vibe Coding Tools
  • AI Design Tools
  • AI Database Tools
  • AI Website Builders
  • AI Testing Tools
  • LLM Evaluations
Follow Us
  • X / Twitter
  • LinkedIn
  • Reddit
  • Discord
  • Threads
  • Bluesky
  • Mastodon
  • YouTube
  • GitHub
  • Instagram
Get Started
  • About
  • Editorial Standards
  • Corrections & Disclosures
  • Community Guidelines
  • Advertise
  • Contact Us
  • Newsletter
  • Submit a Tool
  • Start a Discussion
  • Write A Blog
  • Share A Build
  • Terms of Service
  • Privacy Policy
Explore with AI
  • ChatGPT
  • Gemini
  • Claude
  • Grok
  • Perplexity
Agent Experience
  • llms.txt
Theme
With AI, Everyone is a Dev. EveryDev.ai © 2026
    1. Home
    2. Developers
    3. JustVugg

    JustVugg

    To democratize access to frontier-class AI models by enabling them to run on consumer-grade hardware through optimized inference systems.

    Visit Website

    At a Glance

    1Tool Listed
    17Products
    8Capabilities
    Discussions
    Reggio Emilia, ItalyHeadquarters
    2026Est.
    1Employee
    Focus Areas
    Local Inference
    AI Infrastructure
    Model Management
    Connect
    Latest News
    Colibri v1.7.0 Released: Qwen3.6 support and performance optimizationsAug 19, 2026
    Critical Security Release: Six memory-safety issues patched in v1.6.2Aug 14, 2026
    Markets
    • AI Researchers
    • Individual Developers
    • Privacy-focused organizations
    • Hardware enthusiasts

    AI Tools by JustVugg

    (1)
    View colibri
    colibri tool icon

    colibri

    Open Source MoE Inference Engine

    Local InferenceAI InfrastructureModel Management

    Discussions

    No discussions yet

    Be the first to start a discussion about JustVugg

    Latest News

    08/19/2026

    Colibri v1.7.0 Released: Qwen3.6 support and performance optimizations

    GitHub Changelog
    08/14/2026

    Critical Security Release: Six memory-safety issues patched in v1.6.2

    GitHub Security Advisories
    08/05/2026

    DeepSeek V4 Flash support landed in v1.5.0

    GitHub Changelog
    07/29/2026

    Kimi K3 (2.8T) and Inkling models now run on Colibri v1.3.0

    GitHub Changelog

    Products & Services

    17
    Colibri (colibrì)
    2026-07-19

    A tiny inference engine for running large MoE models (744B to 2.8T) on consumer hardware by treating storage, RAM, and VRAM as a single hierarchy.

    mnem

    Deterministic, dependency-free memory for AI agents where memory lives in a Markdown file.

    nalo

    Open-source Durable Objects for live applications.

    nanoeuler

    GPT-2-style LLM built from scratch in C/CUDA with hand-written backprop and FlashAttention.

    Market Position

    Positions itself as a leaner, MoE-specialized alternative to llama.cpp and vLLM, specifically targeting models that are significantly larger than the available RAM.

    Leadership

    Founders

    VF

    Vincenzo Fornaro

    AI Engineer and high-performance systems developer. Creator of the Colibri inference engine. Previously AI Engineer at SELEA.

    Executive Team

    VF

    Vincenzo Fornaro

    Founder & Lead Developer

    Expert in AI inference, computer vision, and systems programming in C, C++, and Go.

    Founding Story

    Started as a solo project by Vincenzo Fornaro on a 12-core laptop with 25GB of RAM to prove that massive models like GLM-5.2 could run locally without expensive datacenters.

    Business Model

    Revenue Model

    Open source (Apache 2.0). Sponsorships and community donations.

    Pricing Tiers

    Community
    Free

    Full access to source code and prebuilt binaries on GitHub.

    Target Markets

    Industries & Segments
    • AI Researchers
    • Individual Developers
    • Privacy-focused organizations
    • Hardware enthusiasts
    Use Cases
    • Running frontier LLMs on consumer hardware
    • Local and private AI inference
    • Systems research for AI performance
    • Accessibility of massive MoE models
    Notable Customers
    • Community of contributors and testers across various hardware configurations.

    Quick Facts

    Headquarters
    Reggio Emilia, Italy
    Founded
    2026
    Employees
    1
    Office Locations
    Reggio Emilia

    History & Milestones

    2026-06

    Colibri engine starts running in production.

    2026-07-19

    Release of Colibri v1.0.0, the first tagged version.

    2026-07-22

    Release of v1.1.0 with AMD GPU support (HIP/ROCm).

    2026-07-29

    Release of v1.3.0 adding support for Kimi K3 and Inkling models.

    2026-08-05

    Release of v1.5.0 adding support for DeepSeek V4 Flash.

    Key Capabilities

    8
    MoE expert streaming from disk (NVMe)
    Three-tier memory hierarchy (VRAM / RAM / Storage)
    MLA attention with 57x smaller KV cache
    Multi-token prediction (MTP) speculative decoding
    Zero-dependency pure C implementation
    OpenAI-compatible API and web dashboard

    Integrations & Partnerships

    Platform Integrations

    • Linux
    • macOS (Metal)
    • Windows 11 (CUDA/Vulkan)
    • OpenAI API
    • Anthropic Messages API
    • Discord

    Key Partnerships

    Supports models from Z.ai (GLM), Moonshot AI (Kimi), Alibaba (Qwen), Allen AI (OLMoE), and Thinking Machines.

    Connect

    Website
    justvugg.github.io/colibri
    GitHub
    JustVugg
    Discord
    MAaKtQRc

    AI Topics

    3

    JustVugg focuses on these topics:

    Local Inference(1)
    AI Infrastructure(1)
    Model Management(1)
    Back to all developersSuggest an edit