Star 历史趋势
数据来源: GitHub API · 生成自 Stargazers.cn
README.md

AI Glasses for Navigation

Lower-cost visual assistance. Shared making knowledge. Community participation.

Project · Timeline · Community · How it works · Design · Run the demo · 中文介绍

Continuous Facets v0.4 — three-quarter studio render

CONTINUOUS FACETS / DESIGN v0.4
An editable, lensless wraparound enclosure with a straighter brow, clipped corners and integrated temple surfaces.

Blender render of the digital enclosure prototype; not a photograph of a manufactured product.

The project

AI Glasses for Navigation explores affordable, wearable visual assistance for blind and visually impaired people. It brings together three parts: a hardware and enclosure project, software that translates observations into guidance, and community outreach that shares how the device can be made.

Development began in early 2025. The initial building and development phase spanned nearly a year, before the project was uploaded to GitHub in December 2025. Community work began in late 2025 and has continued alongside technical iteration. The project therefore predates this repository: GitHub records the published work, while the earlier building process and community activities form part of the broader project history.

The aim is to make visual-assistance hardware more affordable and its construction more understandable. Sharing the making process matters alongside the device itself: participants can learn how it works and build local knowledge that can continue beyond a single visit or donation.

Today, this repository contains a runnable guidance demo, an editable v0.4 enclosure and its render gallery, plus documentation of community work and hardware costs. Live camera perception, speech output and a fully integrated wearable remain separate development work in the current release.

Project timeline

PeriodMilestoneWhat changed
Early 2025Building beganInitial development of the AI glasses project started, before publication on GitHub.
Throughout 2025Nearly a year of initial developmentBuilding and development continued across the year, forming the early prototype and project direction.
Late 2025Community work beganThe project moved into community outreach and practical explanation, beginning an ongoing effort to broaden participation and share making knowledge.
December 2025First GitHub publicationThe initial repository commit was recorded on December 9, 2025 (UTC). Early documentation described tactile-paving navigation, crosswalk assistance, object search and voice interaction.
July 29, 2026Reproducible software prototypeA hardware-free FastAPI core, interactive dashboard, JSONL replay and in-memory event history were added. An authenticated device-observation gateway, per-device rate limiting, packaging and CI test configuration followed in the same iteration.
September 18, 2026Continuous Facets v0.4The lightweight enclosure direction was published with editable Blender files, GLB/STL exports, geometry checks, multiple rendered views and clearer community and cost documentation.
Late 2025–presentContinuing community engagementOutreach has reached four communities and nearly 60 blind or visually impaired people; three communities received practical explanations of how to make the device. Technical development and community engagement continue in parallel.

The early development dates and community history are provided by the project creator. Publication and software milestones are documented in the commit history. The nearly one-year period refers to initial development during 2025; subsequent iteration continues beyond that period.

Community first

A presenter speaking to seated participants during an in-person session

4 communitiesNearly 60 people3 communities
Engaged through outreachBlind or visually impaired people reachedReceived practical making instruction

Beginning in late 2025, the project expanded from building the device to sharing it through community engagement. The project creator reports reaching nearly 60 blind or visually impaired people across four communities and providing explanations of how to make the device in three communities.

The emphasis is on affordable hardware and knowledge that communities can retain. Making instruction is intended to help local participants continue without relying on the creator for ongoing funding or material donations. Community engagement remains an ongoing part of the project as its reach develops.

These figures describe contact and instruction, rather than devices delivered or daily active users. Independent production volume, long-term use and improvements in mobility have not been quantified. The community work concerns the broader project and earlier prototype; it does not establish deployment of the new v0.4 enclosure.

More on community work and reporting scope →

What the glasses are intended to help with

The broader project explores four everyday assistance workflows. These informed the earlier prototype documentation and remain the context for the current work.

WorkflowIntended assistanceCurrent repository scope
Tactile-paving navigationRecognize a path, describe alignment and identify nearby obstacles.Earlier materials describe YOLO segmentation and path-guidance experiments. The current demo handles obstacle observations; it does not implement a complete tactile-path navigation pipeline.
Crosswalk and traffic-light awarenessIdentify a crosswalk and communicate traffic-light observations with uncertainty.The guidance engine handles structured crosswalk and red/green/unknown light inputs. Live recognition accuracy and safe crossing have not been established by this demo.
Object searchFind a named object and provide cues about its location.Earlier materials describe YOLO-E detection, tracking and MediaPipe hand cues. These are not integrated into the current demo.
Voice interactionAsk questions and receive spoken assistance without relying on a visual display.Earlier materials reference speech recognition and multimodal services. The current core returns text guidance; end-to-end audio integration remains outstanding.

How the system works

The intended wearable pipeline starts with a camera, turns visual input into observations, applies guidance rules and communicates the result through audio. The current runnable software implements the observation-to-guidance portion, making its behavior inspectable without camera hardware or cloud access.

flowchart TD
    A[Demo scenarios or recorded JSONL observations] --> C[Validate structured observation]
    B[Authenticated device observation gateway] --> C
    C --> D[Guidance engine]
    D --> E[Confidence checks and repeated-message suppression]
    E --> F[Text response through dashboard, REST or WebSocket]
    E --> G[Temporary in-memory event history]
    H[Future camera and perception integration] -.-> B
    F -.-> I[Future audio output integration]

Guidance behavior

Observations include a category and confidence value, plus fields such as traffic-light state, obstacle distance or object label. The engine then applies deterministic rules:

  • Traffic lights: a red-light observation produces a stop-and-wait message. A green-light observation is informational and non-actionable: the engine cannot decide whether crossing is possible. An unknown state produces an uncertainty message. Informational green-light messages retain repeated-message suppression.
  • Obstacles: an observation at 1.5 meters or less produces a nearby-obstacle warning; other obstacle observations produce a general caution.
  • Crosswalks: a crosswalk observation produces a message about orientation and checking the signal.
  • Low confidence: observations below the default 0.70 confidence threshold produce an uncertainty message marked as non-actionable.
  • Repeated messages: the same guidance category is suppressed within the default three-second cooldown.
  • Invalid distances: negative or non-finite obstacle distances produce a non-actionable status message, including observations loaded directly from replay files.

These rules respond to supplied observations; they do not prove that a camera detected the scene correctly. Current guidance messages and the demo interface are in Chinese. The project documentation is primarily in English.

Software and hardware responsibilities

ComponentRole and status
FastAPI serviceValidates observations and exposes the demo dashboard, REST endpoints and WebSocket interface.
Guidance engineConverts traffic-light, obstacle and crosswalk observations into text guidance, with confidence and repetition handling.
Replay toolRuns recorded JSONL observations through the same guidance rules for reproducible inspection.
Device gatewayAccepts observations with a matching token in hardware mode; defaults to 120 requests per minute per device. Intended for controlled development networks.
Camera and perceptionEarlier work references ESP32 capture, YOLO/YOLO-E and MediaPipe. Live perception adapters remain integration work for the current core.
Audio and interactionEarlier work references microphones, speakers and cloud speech/multimodal services. These are separate from the current text-response demo.
Power and enclosureThe design provides an enclosure study. Measured electronic-part dimensions, power requirements, fastening and physical assembly still require verification.

Setting hardware mode enables the observation endpoint; it does not automatically connect an ESP32, run a vision model or produce speech. Architecture · Device gateway · Project overview

Data handling

The demo stores the latest 100 guidance events in process memory: timestamp, observation category, confidence, severity, actionability and message text. Restarting the service clears that history. This event store does not retain raw frames, audio, location or device IDs; device IDs are used separately for in-memory rate limiting. Future camera, recording or cloud integrations have separate data-handling requirements. Data boundaries →

Wearable design — Continuous Facets v0.4

The current enclosure is a lensless wraparound design with a straighter brow, clipped corners and tightened temples. Its continuous surfaces come from the housing itself. Six original exterior skin patches were replaced, while the original assembly remains available in a separate reference collection.

Front view of v0.4 with the camera aperture and nose supportsSide view showing the continuous faceted temple
Front / integrated browSide / continuous surfaces
Top view of the U-shaped lensless housingRear three-quarter view of the housing
Top / wraparound layoutRear / interior access

The lower body and upper cover are editable parts. The package includes a Blender assembly, a preserved source reference, GLB preview, shell STL exports, construction scripts and reproducible render scenes. The initial concept board was AI-generated; the published product views are rendered from the editable Blender geometry.

Geometry checks cover shell topology, surface intersections and containment against the protected source geometry. They do not establish calibrated physical units, actual battery/PCB/cable fit, thermal performance, fastening or assembly tolerances. The v0.4 enclosure is a digital design prototype; its STL exports are not production-validated printing files.

Open the model and inspect the checks →

Design in context

v0.4 on a staged everyday desk

v0.4 workbench render with generic precision tools

These are synthetic Blender scenes of the same v0.4 model, rather than records of device use or community activities. The tools are generic props, not supplied accessories or a brand partnership.

All 11 views and image provenance · Editable presentation scenes

Affordability by design

Reducing hardware cost is central to the project, alongside sharing the knowledge needed to make it.

Build contextHardware-only costBasis
Earlier prototype built in ChinaUnder US$20 per unitThe creator's reported cost for the earlier build.
Estimated US buildUnder US$30 per unitThe creator's estimate; not yet verified against a supplier quotation.
New v0.4 enclosureNot separately costedDigital design awaiting fabrication and assembly validation.

These amounts exclude a phone or computer, cloud services, tools, labor and other non-hardware expenses. The earlier prototype's cost is not a verified cost for manufacturing v0.4. A fully itemized bill of materials, including sourcing, shipping, taxes and enclosure fabrication, is not yet available. Cost details →

Current status and continuing work

The repository provides a working software demonstration and inspectable design assets. It does not yet establish a fully integrated, field-validated navigation device.

AreaAvailable nowContinuing work
SoftwareGuidance rules, dashboard, REST/WebSocket interfaces, replay and device ingestion.Live camera/perception adapters and audio output.
EnclosureEditable v0.4 geometry, exports, render scenes and source-geometry checks.Physical scale calibration, measured components, mounting, cable routing, fabrication and fit.
EvaluationRepeatable structured-observation scenarios and repository tests.Video-based perception evaluation, latency and error measurements, message comprehension and physical-prototype testing.
CommunityOutreach across four communities and making instruction in three, reaching nearly 60 people.Continued engagement and documentation of longer-term participation, independent making and use.

Development roadmap →

Run the demo

Use Python 3.11 or later for the current source. Demo mode requires no camera, ESP32, model weights or cloud keys.

git clone https://github.com/DresdenGman/AIGlasses_for_navigation.git
cd AIGlasses_for_navigation
python -m venv .venv
source .venv/bin/activate  # Windows: .venv\Scripts\activate
python -m pip install -r requirements.txt
cp .env.example .env
python main.py

Open the local dashboard or interactive API docs. The dashboard sends sample observations through the guidance engine.

Replay the included observations or run the software tests:

python -m aiglasses.replay demo/events.jsonl
python -m pip install -e '.[dev]'
python -m pytest

The current entry point is main.py, with implementation under aiglasses/. Earlier prototype instructions referring to other entry points describe a different stage of the project. Demo guide →

Explore the repository

AreaFiles and documentation
Project scopeOverview · Roadmap
SoftwareArchitecture · Demo · Device gateway · Data boundaries
Designv0.4 model and downloads · Geometry provenance · Release notes
ImagesRender gallery · Editable presentation scenes
Community and costReach, instruction and hardware costs
ContributionsContributor guide · Issues
中文中文项目介绍

Attribution & responsible use

Earlier project documentation credited AI-FanGe / OpenAIglasses_for_Navigation as an upstream code project. Historical vision and voice work includes upstream contributions; attribution and existing license notices are retained. See the MIT license.

This is an assistive-technology research prototype, not a certified navigation aid. It must not replace a white cane, guide dog, professional mobility support, personal judgment or traffic rules. Outreach figures describe community engagement; the current release makes no measured mobility-performance or field-safety claim.

关于 About

AI Intelligent Blind Glasses System

语言 Languages

Python92.7%
HTML3.7%
JavaScript3.7%

提交活跃度 Commit Activity

代码提交热力图
过去 52 周的开发活跃度
29
Total Commits
峰值: 11次/周
Less
More

核心贡献者 Contributors