LYiHub GitHub avatar

pub-local-jarvis

LYiHub

AI Jarvis is a local full-duplex desktop assistant that continuously understands your screen and system audio via local multimodal models, and provides timely help through pet bubbles, game overlays, or course notes.

Stars

128

7-day growth

No data

Forks

29

Open issues

3

License

MIT

Last updated

2026-07-24

AI repository intelligence
FR-AI / ANALYSIS

Why it is worth attention

It combines screen and audio understanding with a cute desktop pet interface, runs entirely locally for privacy, and offers unique use cases like gaming companion and course note generation without requiring cloud connectivity.

Who it is for

  • Windows 10/11 users who want a local AI desktop assistant
  • Gamers seeking real-time game context hints via transparent overlays
  • Online course learners who need auto-generated notes
  • Developers and AI enthusiasts interested in local multimodal inference

Use cases

  • Desktop companion that gives contextual reminders based on screen activity
  • Gaming companion that provides non-intrusive hints and interactions
  • Auto-generate Markdown course notes with key frames and summaries after online classes
  • Privacy-conscious screen and audio analysis without uploading data to the cloud

Strengths

  • Full local inference with MiniCPM-o 4.5 model, no cloud dependency for core features
  • Active dialog via Ctrl+M keybinding, combined with real-time screen context
  • Transparent game overlay for companion functionality that does not interfere with gameplay
  • Visual daily schedule generation (when configured) and local memory of activity timeline

Considerations

  • Windows 10/11 only, requires x64 processor with AVX2 and at least 12 GiB disk space
  • GPU acceleration recommended; CPU fallback can be very slow (tens of seconds to minutes)
  • Currently text-only interaction; no voice output or real-time voice conversation

README quick start

快速开始

Description

Windows 本地多模态 AI 桌面桌宠,支持屏幕与音频感知。

Related repositories

Similar projects matched by category, topics, and programming language.

gavamedia
Featured
gavamedia GitHub avatar

deltafin

Deltafin is a research project that runs the 2.8-trillion-parameter Mixture-of-Experts model Kimi K3 on a single Apple Silicon Mac (e.g., M1 Max with 64 GB) at about 16 seconds per token, using exact, reproducible inference with local or streaming expert loading.

AI & Machine LearningLarge Language Models
304
lopopolo
Featured
lopopolo GitHub avatar

harness-engineering

Harness Engineering is a methodology for improving coding agent outputs by carefully crafting the environment around them—providing curated context, tools, and executable constraints that encode an organization’s nonfunctional requirements and cumulative lessons.

AI & Machine LearningAI Agents
2,390
slvDev
Featured
slvDev GitHub avatar

esp32-ai

A 28.9 million parameter language model runs on an $8 ESP32-S3 microcontroller entirely on-device, generating simple stories at about 9.5 tokens per second.

AI & Machine LearningLarge Language Models
1,960