Open-Awesome
CategoriesAlternativesStacksSelf-HostedExplore
Open-Awesome

© 2026 Open-Awesome. Curated for the developer elite.

TermsPrivacyAboutGitHubRSS
  1. Home
  2. Godot
  3. NobodyWho

NobodyWho

EUPL-1.2Rustnobodywho-react-native-v5.0.0

A library for running LLMs locally and efficiently on any device with support for Python, Flutter, and Godot.

Visit WebsiteGitHubGitHub
1.5k stars92 forks0 contributors

What is NobodyWho?

NobodyWho is a library that allows developers to run large language models (LLMs) locally and efficiently on any device. It solves the problem of dependency on cloud-based AI services by providing offline, private, and cost-free inference capabilities. The library supports multiple programming environments including Python, Flutter, and Godot, making it versatile for various application types.

Target Audience

Developers building applications that require local AI inference, such as mobile app creators using Flutter, game developers using Godot, and Python developers needing embedded LLM capabilities.

Value Proposition

Developers choose NobodyWho for its guaranteed perfect tool calling, infinite conversation length support, and ability to deploy optimized native code across multiple platforms without licensing fees. Its integration with llama.cpp ensures compatibility with a wide range of GGUF models.

Overview

NobodyWho is an inference engine that lets you run LLMs locally and efficiently on any device.

Use Cases

Best For

  • Building Flutter apps with local LLM capabilities
  • Creating Godot games with integrated AI dialogue systems
  • Developing Python applications that require offline AI inference
  • Implementing reliable tool calling in local AI projects
  • Deploying LLMs on edge devices like Android smartphones
  • Running privacy-sensitive AI applications without cloud dependencies

Not Ideal For

  • Projects targeting iOS or web platforms, as these are not currently supported (issues #114 and #111)
  • Teams needing cloud-based AI for scalable, serverless deployment with minimal local setup
  • Developers requiring support for non-GGUF model formats like PyTorch or ONNX
  • Applications demanding real-time inference on resource-constrained devices without GPU access

Pros & Cons

Pros

Offline Privacy & Cost Savings

Runs LLMs locally without API calls, ensuring data privacy and eliminating ongoing costs, as emphasized in the README's local execution focus.

Reliable Tool Calling

Automatically derives grammars from function signatures for guaranteed perfect tool integration, simplifying development without manual configuration.

Infinite Conversation Length

Uses conversation-aware preemptive context shifting to prevent memory loss in long dialogues, enabling seamless extended interactions.

Cross-Platform Deployment

Supports optimized native code for Windows, Linux, macOS, and Android, allowing deployment on diverse devices from the same codebase.

GPU Acceleration

Leverages Vulkan or Metal for super-fast GPU-powered inference, improving performance for resource-intensive models as highlighted in the features.

Cons

Limited Platform Support

Lacks current support for iOS and web exports, with issues #114 and #111 acknowledging these as future work, restricting its use in some environments.

Model Format Dependency

Restricted to GGUF format models via llama.cpp, requiring conversion for other formats, which can add overhead and limit model selection flexibility.

Setup and Management Overhead

Requires manual download and local management of model files, unlike cloud services with instant access, adding complexity to deployment and updates.

Performance Hardware Reliance

Inference speed is constrained by local GPU availability and hardware specs, which may not match the scalability of cloud-based solutions.

Frequently Asked Questions

Quick Stats

Stars1,522
Forks92
Contributors0
Open Issues6
Last commit1 day ago
CreatedSince 2024

Tags

#gguf#llama-cpp#llm#python#offline-ai#godot-plugin#cross-platform#llm-inference#flutter#godot4#godot#local-ai#tool-calling#godot-engine#inference-engine

Built With

V
Vulkan
l
llama_cpp
M
Metal

Links & Resources

Website

Included in

Machine Learning72.2kGodot9.7k
Auto-fetched 1 day ago

Related Projects

HuggingFace TransformersHuggingFace Transformers

🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.

Stars166,987
Forks34,757
Last commit1 day ago
jiebajieba

结巴中文分词

Stars35,179
Forks6,676
Last commit2 years ago
spacyspacy

💫 Industrial-strength Natural Language Processing (NLP) in Python

Stars33,940
Forks4,731
Last commit6 days ago
HaystackHaystack

Open-source AI orchestration framework for building context-engineered, production-ready LLM applications. Design modular pipelines and agent workflows with explicit control over retrieval, routing, memory, and generation. Built for scalable agents, RAG, multimodal applications, semantic search, and conversational systems.

Stars26,678
Forks3,242
Last commit1 day ago
Community-curated · Updated weekly · 100% open source

Found a gem we're missing?

Open-Awesome is built by the community, for the community. Submit a project, suggest an awesome list, or help improve the catalog on GitHub.

Submit a projectStar on GitHub