CoolFace
Apppublic

dvinix/navora

sourceHugging Facemitupdated 6mo agoView on Hugging Face
0likes
App README

๐Ÿงญ Navora - AI Navigation Assistant for the Blind

Real-time assistive vision system providing navigation guidance for blind and visually impaired users.

โœจ Features

AI-Powered Detection

  • โ€”YOLOv8 โ€” Real-time object detection
  • โ€”BLIP-2 โ€” Scene understanding and description
  • โ€”MiDaS โ€” Depth estimation (GPU only)

Accessibility Features

  • โ€”๐ŸŽฏ Visual Feedback โ€” Thin bounding boxes around detected objects
  • โ€”๐Ÿ—ฃ๏ธ Voice Guidance โ€” Clear audio navigation instructions
  • โ€”๐Ÿ“ข Scene Description โ€” Automatic description when standing still
  • โ€”๐Ÿ“ณ Haptic Feedback โ€” Vibration patterns for different obstacles
  • โ€”๐Ÿšจ Smart Alerts โ€” Context-aware, non-repetitive announcements

๐Ÿš€ How to Use

On Mobile Device

  1. 1.Open this Space on your mobile browser
  2. 2.Allow camera access (uses back camera)
  3. 3.Point camera at your path
  4. 4.Receive real-time guidance:
  5. 5.Voice: Spoken navigation instructions
  6. 6.Vibration: Tactile feedback for obstacles
  7. 7.Visual: Bounding boxes (for sighted assistants)

Navigation Actions

  • โ€”Forward โฌ†๏ธ โ€” Path is clear, move forward
  • โ€”Stop ๐Ÿ›‘ โ€” Obstacle directly ahead
  • โ€”Left โฌ…๏ธ โ€” Obstacle on right, move left
  • โ€”Right โžก๏ธ โ€” Obstacle on left, move right

Scene Description

  • โ€”Stand still for 5 seconds
  • โ€”System automatically describes surroundings
  • โ€”Example: "5 objects detected: 2 persons, 1 car, 1 bicycle"

๐ŸŽฏ Obstacle Detection

The system prioritizes obstacles based on:

  • โ€”Position โ€” Center obstacles are highest priority
  • โ€”Confidence โ€” Higher confidence = more reliable
  • โ€”Type โ€” People, vehicles prioritized over static objects
  • โ€”Size โ€” Larger objects get more attention

๐Ÿ”ง Technical Details

Architecture

Camera โ†’ YOLOv8 Detection โ†’ Priority Analysis โ†’ Navigation Guidance
                                                โ†“
                                    Voice + Vibration + Visual

Performance

  • โ€”Inference: ~1-2 seconds per frame (CPU)
  • โ€”Frame Rate: ~1-2 FPS (optimized for battery)
  • โ€”Model Size: 6MB (YOLOv8n)
  • โ€”Accuracy: 45%+ confidence threshold

Optimizations

  • โ€”CPU-optimized model selection
  • โ€”Adaptive frame rate
  • โ€”Smart alert deduplication
  • โ€”Low battery mode support

๐Ÿ“ฑ Mobile Compatibility

Works on:

  • โ€”โœ… iOS Safari (iPhone/iPad)
  • โ€”โœ… Android Chrome
  • โ€”โœ… Android Firefox
  • โ€”โš ๏ธ Requires HTTPS for camera access

๐Ÿ”’ Privacy

  • โ€”All processing happens on server
  • โ€”No video storage
  • โ€”No user tracking
  • โ€”Session data cleared after 2 minutes

๐Ÿšง Limitations

  • โ€”Requires internet connection
  • โ€”Works best in good lighting
  • โ€”May struggle with:
  • โ€”Very dark environments
  • โ€”Fast movement
  • โ€”Very crowded scenes
  • โ€”Reflective surfaces

๐Ÿ”ฎ Future Enhancements

  • โ€”๐Ÿ“ฑ Mobile app with offline capability
  • โ€”๐ŸŒ™ Night vision mode
  • โ€”๐Ÿ—บ๏ธ Landmark recognition
  • โ€”๐Ÿšจ Emergency SOS feature
  • โ€”๐ŸŽค Voice commands
  • โ€”๐Ÿ”Š Spatial audio (3D sound)

๐Ÿ“š Documentation

  • โ€”Accessibility Features
  • โ€”Implementation Guide
  • โ€”Mobile Edge Deployment

๐Ÿค Contributing

Feedback from blind users is invaluable! Please test and share your experience.

๐Ÿ“„ License

MIT License - Free to use and modify