Browse applications built on huggingface's technologies. Explore PoC and MVP applications created by our community and discover innovative use cases for huggingface's technologies.
Aura is a sovereign, neuro-symbolic AI operating system that turns natural-language intent into transparent, governed action through local parsing, agent collaboration, bounded experimentation, verifiable learning, and human-controlled decision gates.
SOAP Copilot turns raw doctor-patient conversations into structured SOAP notes, ICD-10 codes, and patient-friendly summaries in seconds, using a 3-agent Llama 3.3 70B pipeline built and fine-tuned on AMD hardware.
Thymus is a lightweight hybrid token-efficient router designed to maximize accuracy while minimizing token costs in multi‑task LLM pipelines. It dynamically routes user queries across local and remote models on LLM providers.
An agentic pipeline that takes any CUDA repo and autonomously ports it to ROCm/HIP — analyze → transpile → compile → test → benchmark on real AMD GPUs — with an LLM repair loop fixing everything the mechanical tools can't.
Verity turns any claim into a traceable verdict. Multi-model consensus checks headlines, stats, and posts against real sources in seconds running live on AMD GPU's via Fireworks AI ,Gemma and Huggingface's APIs, open and inspectable, not a black box.
Contxt is a portable, privacy-first context layer for every AI. It remembers you across Claude, ChatGPT and Gemini and injects your context live, while an on-device Gemma gateway keeps your private data encrypted so it never leaves your device.
Zero-trust enterprise privacy middleware for LLMs. AegisLayer uses AMD ROCm-accelerated NER to redact sensitive PII locally before cloud inference, and seamlessly restores original values in the response using an ephemeral in-memory vault.
A local-first AI routing agent that uses a fine-tuned DistilBERT classifier to choose cheap or stronger Fireworks models, reducing routing token cost while preserving accuracy across eight task categories.
An AMD-powered HS-code classifier that turns a plain-English product description into its tariff code, import duty, and licensing restrictions — every number backed by an official government source URL.
SchemaIntern converts natural language questions into accurate SQL queries by understanding your database schema. Connect any PostgreSQL database, ask questions in plain English, and get safe, optimized SQL instantly (no SQL knowledge required.)
Turns NASA's Surya solar foundation model into threat forecasts on AMD GPUs, routes them through a 5-agent system (via Fireworks AI) that assesses risk and issues real satellite, grid, and comms protective actions visualized in a live 3D dashboard.
Hybrid AI router that intelligently routes user queries between deterministic tools, local OpenVINO models, and Fireworks AI, optimizing for accuracy, latency, and cost while minimizing unnecessary remote API calls.
CredIQ is a real-time credit card spending tracker with a live budget bar, AI chatbot, photo-based card analyzer, and offer recommendations — built on AMD infrastructure via Fireworks AI and Firebase.
Smart-Router Agent answers all 8 Track 1 task categories at a fraction of the token cost. A free on-device model handles the easy tasks; the paid Fireworks API is reserved for hard math, logic, and code.
A typed, decaying, graph-aware memory layer for software engineers that runs fully local on AMD ROCm / Instinct GPUs — no cloud, no data leaves the machine.
A hybrid, token-efficient AI agent for 8 task categories. Zero-token heuristic routing sends easy tasks to a bundled local model (free) and hard ones to the best Fireworks model, clearing the accuracy gate while minimising billed tokens.
ReachIQ AI's Video RAG engine ingests competitor YouTube videos, builds a vector knowledge base, and generates strategic competitive intelligence reports — powered by Gemma 3 4B running natively on AMD Developer Cloud via ROCm + vLLM.
FlowForge turns plain English into ready to import n8n automation workflows. We fine-tuned Gemma 3 4B on an AMD Radeon PRO W7900 using LoRA and real n8n templates taking the model from generating workflows to twenty percent fully valid schema JSON.
Clinical decision support for elderly polypharmacy. A deterministic engine 102 drug interactions,61 AGS Beers rules, 58 STOPP/START v3 criteria runs in ~3ms on every request, gated on age and eGFR, escalating to a cloud LLM only when the case warrants it.
A neural network that grows its own brain mid-training. When loss spikes, a local Gemma-2-2b-it LLM writes, validates, and hot-swaps in a new expert module — no human, no restart. Verified on real AMD ROCm hardware.
Autonomous AI agent that ports CUDA to AMD ROCm, validated on MI300X. Gemma 4 12B + HIPIFY-first. AMD Hackathon ACT II.
A local DistilBERT binary router sends each query to the cheapest suitable Fireworks model, preserving answer quality while using zero routing tokens across eight task categories.
VISTA generates accurate, scene-aware video captions by combining visual and audio understanding. It analyzes videos scene by scene and produces four unique caption styles tailored to different audiences.
GlazeSmith is a physics-grounded digital twin for ceramic glazes. It predicts how a recipe fires - crazing risk, surface, color - then renders the result live on a rotatable 3D vase, verified end-to-end. Classifiers trained on AMD ROCm.
A dual-model DevSecOps Chaos Engine. Preflight AI mathematically maps Terraform architectures and offloads massive Monte Carlo failure simulations to an AMD MI300X GPU to prevent catastrophic cloud outages before the code is merged.
I love engineering stochastic multi-agent execution graphs, optimizing bare-metal hardware primitives over non-uniform memory access distributed compute structures
Privacy-first AI coding assistant powered by AMD ROCm and Fireworks AI, delivering lightning-fast local inference, intelligent cloud routing, real-time code analysis, bug detection, and personalized developer guidance.
Local-first agent that answers every task on-box at ZERO Fireworks tokens: deterministic solvers (safe-AST math, lexicon sentiment, code execution) plus a bundled Qwen2.5-1.5B CPU model with grammar-constrained decoding and tight token caps.
Hybrid token-efficient AI agent that analyzes prompt complexity and routes requests to either a local Qwen model or a Fireworks-hosted LLM, reducing API usage while preserving response quality and explainable decision-making.
CaptionCraft AI transforms videos into contextual captions through a multi-stage AI pipeline. It extracts keyframes and audio, transcribes speech, generates scene-aware summaries, and creates captions using Gemini, Fireworks AI, or local LLMs.