Build and own your own AI using just your iPhone and iPad

A plan for Justin Chance · Prices checked October 7, 2026 (PT) · Research only: no accounts created, nothing bought, no app code changed.

The short version

Yes, you can do all of this from your iPhone 16 Plus and iPad. The heavy work runs on computers you rent by the hour in the cloud. You use them through a web browser, and through the Linux box that already runs Chance AI.

Quick terms. Model: the AI “brain” file. Open‑weight: you can download that file and keep it. GPU: the chip that runs AI. VRAM (GB): the GPU's memory; bigger models need more. LoRA: a small add‑on file that teaches a model your style or knowledge without retraining the whole thing. Endpoint: the web address your app sends questions to.

1 · Pick a base model you're allowed to own

You don't build the brain from nothing. You start from a free, open‑weight model whose license lets you modify it and use it commercially, then you shape it.

Model familyLicenseGood for you?Catch
Qwen3.8‑27B / Qwen3.6‑27B (Alibaba)Apache 2.0Best pick. Strong at code, and runs on a single rented GPU.Qwen's top “Max/Plus” models are only available through Alibaba's paid service; you can't download them. Check each model's license file.
Gemma 4 (Google, up to 31B)Apache 2.0 (new with Gemma 4)Great backup. Similar size.Older Gemma 1–3 models use Google's stricter custom terms, so make sure the file you download says “Gemma 4”.
Mistral Small 4Apache 2.0Good, smaller option.The larger Devstral 2 uses a “Modified MIT” license: companies earning more than $20M a month need a paid license. Some Mistral models aren't downloadable at all.
DeepSeek V4 Flash / ProMIT (very permissive)Fine license, but too big to start with.Flash has 284B parameters and Pro has 1.6T, so you'd need several large GPUs at once. Expensive.
Llama 4 (Meta)Llama 4 Community License (not truly open source)Usable, but has the most strings attached.If you share it, you must show “Built with Llama” and start the model's name with “Llama”. You must follow Meta's use policy. Services with more than 700M users need Meta's permission. The multimodal (image‑reading) rights don't apply to people based in the EU.

Plain meaning: under Apache 2.0 or MIT, the copy you download and the changes you make are yours to keep and use forever. Meta, Google, or Alibaba can't later turn off a file you already have. Your only duty is to keep their license notice with it.

2 · Renting GPU power from your phone (pay by the hour)

All of these run in a phone browser. “24 GB” is enough for small and medium models. “80 GB” is for bigger models and faster training.

Host24 GB GPU80 GB GPUUp‑front moneyStorage fees / catches
RunPod (pods)RTX 4090: $0.34/hr (community) · $0.74 (secure)
RTX 3090: $0.22 / $0.50
A100 80GB: $1.19–$1.39 (community) · $1.59 (secure)
H100: $1.99–$2.89
$10 minimum, prepaid and non‑refundable. You need at least 1 hour of credit to start a machine.Disk: $0.10/GB/mo while running, but $0.20/GB/mo while stopped. Network storage: $0.07/GB/mo. Billed per second. 48 GB A40: $0.35 / $0.49.
Vast.ai (marketplace)RTX 4090 from about $0.36/hr · RTX 3090 about $0.14A100 SXM4 from about $0.37 · H100 about $1.73$5 minimum depositPrices change constantly; these are the lowest live offers on Oct 7. The machines belong to many different hosts, so they're less private (filter for verified datacenters). A stopped machine still bills for storage until you destroy it.
LambdaA10: $1.29/hrH100 PCIe: $3.29/hrNo prepay: card on file, billed weekly by the minuteStorage: $0.20/GB/mo, billed even when nothing is using it. Pricier, but simple.
ModalL4: $0.80/hrA100 80GB: $2.50/hr · H100: $3.95/hr$0 “Starter” plan that includes $30/mo of compute (not a subscription)You set it up with code, so you'd do it from the Linux box, not by tapping. It shuts down when idle, so you pay nothing when it isn't in use. Storage: $0.09/GB/mo.
Hugging Face (Endpoints)L4: $0.80/hrA100 80GB: $2.50/hr · H100: $4.50/hrPay as you go (skip the PRO subscription)Free accounts get 100 GB of private storage, which is a good place to keep your model files.
Together AI (fine‑tuning service)Charges per amount of training text, not per hour (see Stage B). On‑demand H100 clusters cost $3.99/hr.$5 minimum prepaid. Credits don't expire.Easiest way to train without managing a machine, and you can download the result.

3 · The three stages, with real costs

Stage A: run an open model privately and add “My AI” to Chance AI

  1. On RunPod, start a pod with a ready‑made vLLM or Ollama template. These are free programs that run a model and give it an OpenAI‑style web address. Use a 48 GB A40 ($0.49/hr) or a 24 GB 4090 ($0.34–$0.74/hr).
  2. Load Qwen3.8‑27B. On a 24 GB card, use a 4‑bit (compressed) copy. Set a password (vLLM's --api-key).
  3. In Chance AI, add a third mode, “My AI”, next to Standard and Private. It points at that address. Your app's Private mode already talks to Tinfoil and Venice in this same “OpenAI‑style” format, so this is a small change the App Builder can make.

Cost: about $1–$3 for a 2–4 hour test. Watch out: leaving it on 24/7 costs about $245–$533 a month ($0.34–$0.74 × 720 hours). So turn it on only when you need it, or use RunPod “Serverless” (24 GB at $0.69/hr, charged only while it's answering, with a slow first reply).

Stage B: customize it with LoRA on your own data and code

Stage C: what “creating your own AI” really means long term

4 · Bringing all your Chance AI code over

Short answer: yes. Everything you've built works with your own AI setup. Chance AI is ordinary code (Python and web files) that you own. It isn't locked into any company's platform.

What's in it now (from a read‑only look at the folder): the Flask server (server.py), agents (agents.py), the App Builder (builder.py), Private‑mode providers (private_providers.py), the cloud‑computer feature (computer.py), extras, voice, the web app (index.html, app.js, app.css), tests, and a data/ folder (settings, agents and their history, analytics). There is also a Replit copy.

  1. Put it in a private GitHub repo. Free accounts get unlimited private repos. Keep passwords and API keys out of the repo: data/ already has a .gitignore (a list of files git skips), so add your key files to it.
  2. Run it anywhere. Keep it on the always‑on Linux box it uses today (cheap, already working with your Cloudflare tunnel), or run it on the same rented GPU machine as your model.
  3. Add “My AI” as a third provider next to Standard (Gemini/OpenRouter) and Private (Tinfoil/Venice). It points at your model's OpenAI‑style address from vLLM or Ollama.
  4. Outside services keep working. Gemini, OpenRouter, Inworld voice, and Tinfoil keep running on your own keys. You can replace them one at a time with your own model whenever you like.
  5. Your app becomes your training data. Your Chance AI chats, agent histories, playbooks, and change requests can be turned into the example conversations Stage B needs. Remove anything private first. Catch: your own messages are fine to use, but some providers limit training on their AI's replies. Google's Gemini API terms say you may not use it “to develop models that compete with” Gemini. Train on your own writing, on replies from open models, or on replies you've corrected yourself.

5 · How your code and model stay yours

6 · Honest note

Rented cloud GPUs are far more powerful than any home computer you'd buy. You can rent a top data‑center chip (an H100) for about $1.99–$3.99/hr, and nothing you could put in a home matches it. But: it needs an internet connection, you pay every minute a machine is on (even if you forget it), and stopped machines still charge for storage. Always delete machines when you're done, and keep deposits small.

Recommended starter path (all from iPhone/iPad)

  1. Today, free: create a private GitHub repo and put the Chance AI code in it, with keys excluded.
  2. First paid step, about $10 deposit (about $1–$3 used): on RunPod, start a 48 GB A40 with the vLLM template, load Qwen3.8‑27B, and try it from Chance AI as “My AI”. Then delete the pod.
  3. Next, about $5–$10 more: export about 1–2M tokens of your own cleaned Chance AI conversations. Run one LoRA fine‑tune (Together, $5 minimum, or on RunPod). Download the result to your private Hugging Face repo.
  4. Then: load your LoRA on top of Qwen in “My AI”. That's your AI: your code, your data, your model files.

Total to get started: about $15–$20, all pay‑as‑you‑go, with no subscriptions.

Sources

Prices change often, especially on Vast.ai, so check the live page before you rent. On RunPod, “Community” means cheaper machines from vetted outside hosts, and “Secure” means higher‑reliability data centers.

↑ Top