mudler/parakeet.cpp
PublicFast and portable Parakeet implementation in C++ with ggml
Find trending repositories by name or description.
Fast and portable Parakeet implementation in C++ with ggml
LLM inference in C/C++
An all-in-one, pure C++ inference engine for audio models, powered by ggml. Supports TTS, STT, VAD, voice conversion, music generation, and more, with highly optimized performance. No Python dependency.
A curated list of awesome C++ (or C) frameworks, libraries, resources, and shiny things. Inspired by awesome-... stuff.
Port of OpenAI's Whisper model in C/C++
Diffusion model(SD,Flux,Wan,Qwen Image,Z-Image,...) inference in pure C/C++
Open-source, bring-your-own-key (BYOK) multi-model AI chat client for iOS, Android and the web. One LLM client for OpenAI, Anthropic, Google Gemini, OpenRouter, DeepSeek and ten more providers, plus any OpenAI-, Anthropic- or Gemini-compatible endpoint — Ollama, LM Studio, llama.cpp, vLLM. Local-first, self-hostable, no account. AGPL-3.0.
Free, open-source, 100% offline voice dictation for Linux. Speak and type anywhere via whisper.cpp, Whisper & VOSK engines, GPU-accelerated, works on X11 + Wayland!
Go With Your Own Intelligence! Use Go for hardware accelerated local inference with llama.cpp, whisper.cpp, and stable-diffusion.cpp directly integrated into your Go applications. Kronk provides a high-level API and production ready model server.
Abseil Common Libraries (C++)
handwritten harness for AI, not vibecoded, everything is a module/plugin, tailor made for use with local models, tiny system prompt (around 1k by default), support for exclusive llamacpp features, strong security, strict zero-trust-in-the-ai policy (the code hard limits the ai, no amount of prompt engineering will bypass it)
~95% on SimpleQA (e.g. Qwen3.6-27B on a 3090). Supports all local and cloud LLMs (llama.cpp, Ollama, Google, ...). 10+ search engines - arXiv, PubMed, your private documents. Everything Local & Encrypted.
A C++ header-only HTTP/HTTPS server and client library
A set of performant RDNA/HIP patches against llama.cpp focusing on high speed inferencing with numerics assurance
Experimental llama.cpp fork for inference research and development
static analysis of C/C++ code
An NMOS (Networked Media Open Specifications) Registry and Node in C++ (IS-04, IS-05)
A Pure Rust based LLM, VLM, VLA, TTS, OCR Inference Engine, powering by Candle & Rust. Alternate to your llama.cpp but much more simpler and cleaner..
lightweight, standalone C++ inference engine for Google's Gemma models.
cpprefjpサイトのMarkdownソース
A native macOS app that allows users to chat with a local LLM that can respond with information from files, folders and websites on your Mac without installing any other software. Powered by llama.cpp.
Pure C++ implementation of several models for real-time chatting on your computer (CPU & GPU)
Lightweight recording and sampling of performance counters for specific code segments directly from your C++ application.
Quickstart template for GDExtension development with Godot
A collection of resources on modern C++
Implementations of the A* algorithm in C++
MNE-CPP: The C++ framework for real-time functional brain imaging.
Source Code Generation for Automatic Differentiation using Operator Overloading
Source code for the book Real-Time C++, by Christopher Kormanyos