Running Text Embedding Models on Jetson Nano for RAG
S_Anv · reddit · 2026-07-30
The poster is asking if anyone has experience running text embedding models (like microsoft/harrier-oss-v1-0.6b) on a Jetson Nano (2GB/4GB). The goal is to build a local RAG (Retrieval-Augmented Generation) system, and they are looking for insights on the actual inference speed and feasibility on this edge hardware.
More from Infra
- 5-7 Year Grid Interconnection Queues May Reset AGI Compute Predictions — Novel-Lifeguard6491 · 2026-07-30
- New S3 Client Delivers 20x Throughput Increase on a Single Core — mgill25 · 2026-07-30
- Microsoft's Fuel-Cell-in-a-Rack Architecture Sparks Debate for Being 'Aggressive' — jwt0625 · 2026-07-30
- Hugging Face Security Report Reveals Spectacular Ops and Monitoring Failure — basedjensen · 2026-07-30
- Mac Inference Speed Surges 80% as Open Source Community Breaks Performance Limits — gajesh · 2026-07-30
- Sentinel Framework Repurposes Old Android Phones into LAN-based AI Vision Nodes — tom_doerr · 2026-07-30