EveryDev.ai
Subscribe
Home
Developers

3,795+ AI companies

  • Radar
  • Trending
AI Tools by Topic
  • AI Coding Assistants
  • Agent Frameworks
  • MCP Servers
  • AI Prompt Tools
  • Vibe Coding Tools
  • AI Design Tools
  • AI Database Tools
  • AI Website Builders
  • AI Testing Tools
  • LLM Evaluations
Follow Us
  • X / Twitter
  • LinkedIn
  • Reddit
  • Discord
  • Threads
  • Bluesky
  • Mastodon
  • YouTube
  • GitHub
  • Instagram
Get Started
  • Users
  • Rate Tools
  • About
  • Editorial Standards
  • Corrections & Disclosures
  • Community Guidelines
  • Advertise
  • Contact Us
  • Newsletter
  • Submit a Tool
  • Start a Discussion
  • Write A Blog
  • Share A Build
  • Terms of Service
  • Privacy Policy
Explore with AI
  • ChatGPT
  • Gemini
  • Claude
  • Grok
  • Perplexity
Agent Experience
  • llms.txt
Theme
With AI, Everyone is a Dev. EveryDev.ai © 2026
    1. Home
    2. Developers
    3. Xuan-Son Nguyen

    Xuan-Son Nguyen

    Xuan-Son Nguyen is an individual software engineer and open-source creator, not a company. He focuses on machine learning, low-level systems, and privacy-preserving/on-device AI; his stated personal motto is “AI for fun, not for profit.”

    Visit Website

    At a Glance

    1Tool Listed
    6Products
    10Capabilities
    Discussions
    FranceHeadquarters
    Focus Areas
    Local Inference
    AI Development Libraries
    AI Infrastructure
    Connect
    Latest News
    wllama 3.8.1 released with decision-model support through the systemone-compat APIOct 2, 2026
    wllama 3.6.1 released with upstream llama.cpp synchronization and removal of pre-built WASM from the repositoryAug 27, 2026
    Markets
    • Open-source developers and JavaScript/TypeScript engineers
    • Researchers and hobbyists experimenting with local and browser-based LLMs
    • Privacy-conscious users who prefer on-device AI
    • Teams building web AI demos and client-side inference applications
    • +1 more

    AI Tools by Xuan-Son Nguyen

    (1)
    View wllama
    wllama tool icon

    wllama

    Browser LLM Inference via WebAssembly

    Local InferenceAI Dev LibrariesAI Infrastructure

    Discussions

    No discussions yet

    Be the first to start a discussion about Xuan-Son Nguyen

    Latest News

    10/02/2026

    wllama 3.8.1 released with decision-model support through the systemone-compat API

    github.com
    08/27/2026

    wllama 3.6.1 released with upstream llama.cpp synchronization and removal of pre-built WASM from the repository

    github.com
    08/16/2026

    wllama 3.6.0 released with updated llama.cpp synchronization, n_parallel support, and partially downloaded-file fixes

    github.com
    05/11/2026

    wllama 3.1 released with WebGPU support and a single pre-built WASM binary for single- and multi-thread use

    github.com

    Products & Services

    6
    wllama / @wllama/wllama
    Active releases through 2 October 2026; the repository identifies version 3.8.1 as latest.

    A WebAssembly binding and TypeScript/JavaScript library for llama.cpp that runs GGUF large-language-model inference directly in the browser without a backend or external API. It is maintained by Xuan-Son Nguyen and distributed as a pre-built npm package, with a Hugging Face demo and documentation.

    Jelly Music
    2014

    An Android application created by Nguyen; it attracted more than 100 users and was later removed from Google Play because of compatibility issues, with source code retained on GitHub.

    Nui Kernel
    2015

    A modified Android/Linux kernel project providing low-level controls including overclocking, underclocking, voltage control, and LED manipulation.

    Chatbot CNH
    2017

    A Facebook Messenger chatbot for anonymous conversations with strangers; it reached more than 200 users on its first day and evolved into bespoke chatbot and web-development work.

    Market Position

    wllama brings llama.cpp to the browser through WebAssembly and WebGPU, emphasizing local execution, privacy, and an OpenAI-compatible interface. The browser-inference alternatives identified in the associated technical material include ONNX Runtime/Transformers.js and WebLLM; wllama differentiates through its llama.cpp/GGUF compatibility, multimodal support, tool calling, model splitting, and worker-based execution.

    Leadership

    Founders

    XN

    Xuan-Son Nguyen

    Software engineer from Vietnam who studied at Université d’Aix-Marseille and INSA Centre Val de Loire in France. His background spans web development, machine learning, cybersecurity, C++, Linux/kernel work, and open-source AI; he worked at Botfuel and Snowpack before joining Hugging Face in August 2024.

    Founding Story

    This is a personal portfolio and independent open-source body of work rather than a formally founded company. Nguyen describes starting with coding and electronics as a child, then building increasingly ambitious software, hardware, web, cybersecurity, and AI projects; his stated motivation is impact and exploration rather than profit.

    Business Model

    Revenue Model

    Nguyen’s work is primarily open source and explicitly described as “AI for fun, not for profit.” The personal site links to GitHub Sponsors and Buy Me a Coffee, indicating optional sponsorship and donations rather than a paid product subscription model.

    Target Markets

    Industries & Segments
    • Open-source developers and JavaScript/TypeScript engineers
    • Researchers and hobbyists experimenting with local and browser-based LLMs
    • Privacy-conscious users who prefer on-device AI
    • Teams building web AI demos and client-side inference applications
    • Users and developers of llama.cpp, GGUF models, Hugging Face, and Ollama ecosystems
    Use Cases
    • Private, local LLM inference in a web browser
    • Browser-based chat, completion, and embedding applications
    • Client-side multimodal AI using image or audio input
    • Web demos and prototypes that need on-device inference without a server
    • Tool-calling and decision-model experiments
    • Developers integrating llama.cpp capabilities into JavaScript or TypeScript applications

    Quick Facts

    Headquarters
    France

    History & Milestones

    May 2026

    Released wllama 3.0 with a llama-server-based architecture, OpenAI-compatible APIs, multimodal inputs, tool calling, and Jinja chat-template parsing, followed by 3.1 with WebGPU support and a unified WASM build.

    2 October 2026

    Released wllama 3.8.1 with decision-model support through the systemone-compat API and an example demo.

    2 February 2025

    Presented “wllama: bringing llama.cpp to the web” at FOSDEM 2025, describing the WebAssembly-based browser inference project.

    August 2024

    Joined Hugging Face after contributing to llama.cpp and being approached by Hugging Face CTO Julien Chaumond about on-device LLM technologies.

    October 2024

    Contributed to Hugging Face’s Ollama compatibility layer, which enabled GGUF models from Hugging Face to be used directly with Ollama; the launch generated more than 20,000 download requests in 24 hours.

    Key Capabilities

    10
    In-browser LLM inference using WebAssembly SIMD, with no backend or GPU required
    WebGPU acceleration
    Multimodal image and audio input
    OpenAI-compatible typed API for chat completions, completions, and embeddings
    Native tool calling
    Jinja-based chat-template parsing

    Integrations & Partnerships

    Platform Integrations

    • npm package @wllama/wllama
    • Hugging Face Space demo at https://huggingface.co/spaces/ngxson/wllama
    • Hugging Face model repositories and GGUF models
    • JavaScript/TypeScript, React, and ES6-module usage
    • llama.cpp and llama-server-compatible APIs
    • Browser WebAssembly and WebGPU, with documented browser compatibility limitations

    Key Partnerships

    Collaboration with Hugging Face on on-device LLM technologies and the Hugging Face-to-Ollama GGUF compatibility layer.
    wllama is built on and integrates the llama.cpp inference project; its WebGPU work is connected to the upstream llama.cpp ecosystem.

    Connect

    Website
    ngxson.com/
    GitHub
    ngxson
    X / Twitter
    ngxson
    LinkedIn
    ngxson
    Bluesky
    ngxson.hf.co

    AI Topics

    3

    Xuan-Son Nguyen focuses on these topics:

    Local Inference(1)
    AI Development Libraries(1)
    AI Infrastructure(1)
    Back to all developersSuggest an edit