- SignalDesk2天前
Original Summary
I’ve been building a project called NodeAI . The basic idea is kind of like GeForce NOW for AI : the phone is the interface, while the actual AI inference runs remotely on a GPU. Current setup: iPhone → NodeAI → Cloudflare → RTX 3090 → Qwen 3.5 9B → streamed response back to the phone The interesting part for me is that the phone doesn't need to have the compute locally. The GPU is essentially a remote AI computer that the phone connects to. I currently have: iPhone-first chat interface authenticated connection remote Qwen inference token-by-token streaming GPU-backed inference on an RTX 3090 direct OpenAI-compatible inference API I built the prototype mainly to see whether I could make remote GPU AI feel like a native phone experience. Demo: https://reddit.com/link/1wit224/video/wx0v2d20u2qh1/player I'm curious what people think of the concept. The long-term idea is to make the compute layer interchangeable so users don't necessarily need to own a powerful computer themselves.   submitted by   /u/bluedream212 [link]   [comments]
- 情报分类:硬件与数码
- 分类依据:内容涉及硬件、数码产品或通信卡
- 信息来源:Reddit · SideProject
- 发布时间:2026/9/17 20:47:32
- 暂无回复