Minimum
12 GB to under 20 GB VRAM
0/1 accepted evidence sets
Keep your existing setup. Connect a worker, choose when it runs, and earn AIPG for completed, accepted jobs.
Work availability and earnings vary. No signup bonus or guaranteed rate.
LM Studio, Ollama, vLLM, SGLang, LMDeploy, or another OpenAI-compatible server. The endpoint matters, not the engine.
v0.3.9 · Release checksums and artifact identity checked.
cd ~/Downloads
chmod +x install-worker.sh
./install-worker.sh
~/.local/bin/grid-inference-worker --verify-runtime
~/.local/bin/grid-inference-workerThe local setup wizard opens at http://localhost:7861.Your model files stay local. Control your schedule, concurrency and pause from the worker.
The worker never needs a wallet private key. Enter credentials only in the local setup wizard, never in public posts or shell commands. Workers process plaintext prompts and outputs.
Need setup help?Operator planner
Choose the workload you want to offer and adjust the example specs to match your machine. Nothing is submitted. Your local backend must still pass a generation test.
Likely starting point
Start Ollama or an OpenAI-compatible local server, load a model that fits your hardware, then select and test it in the worker. VRAM alone cannot determine model fit: quantization and context length also matter.
A verified worker download is available for your selected platform. Check current network needs separately; they are not recommendations for what fits your GPU.
View worker downloadsNetwork-priority text route
gpt-oss-120b
1 serving worker and 2,385 jobs producing 1,216,939.63 accepted den in 30 days. Priority uses accepted den and missing replicas, not raw request count or hardware compatibility. Advertise it only when your backend genuinely serves that model.
Live network opportunity
Jobs per worker is a rough workload-share signal, not a payout forecast. Throughput is observed across the existing Grid and is not a benchmark for your GPU.
Smollm-135m
text
deepseek-v4-flash-nvfp4
text
gpt-oss-120b
text
qwen3-27b
text
qwen38-flash-next-125b-nvfp4
text
FLUX.2 Klein 4B FP8
image
gpt-oss-20b
text
LTX-2.3
video
ace-step-v1.5-xl-turbo
audio
z-image-turbo
image
Krea 2 Turbo
image
LTX Director 2.0
video
LTX-2.3 Audio
video
| Model | Type | Workers | Jobs, 30d | Jobs / worker | Observed | Resilience |
|---|---|---|---|---|---|---|
| Smollm-135m | text | 1 | 77,438 | 77438.0 | 396.6 tok/s | Single-worker risk |
| deepseek-v4-flash-nvfp4 | text | 1 | 3,008 | 3008.0 | 84.6 tok/s | Single-worker risk |
| gpt-oss-120b | text | 1 | 2,385 | 2385.0 | 53.0 tok/s | Single-worker risk |
| qwen3-27b | text | 1 | 757 | 757.0 | 75.3 tok/s | Single-worker risk |
| qwen38-flash-next-125b-nvfp4 | text | 1 | 592 | 592.0 | 6.6 tok/s | Single-worker risk |
| FLUX.2 Klein 4B FP8 | image | 1 | 109 | 109.0 | 3.7s average | Single-worker risk |
| gpt-oss-20b | text | 1 | 103 | 103.0 | 128.3 tok/s | Single-worker risk |
| LTX-2.3 | video | 1 | 48 | 48.0 | 65.9s average | Single-worker risk |
| ace-step-v1.5-xl-turbo | audio | 1 | 28 | 28.0 | 13.1s average | Single-worker risk |
| z-image-turbo | image | 1 | 26 | 26.0 | 17.5s average | Single-worker risk |
| Krea 2 Turbo | image | 1 | 23 | 23.0 | 8.3s average | Single-worker risk |
| LTX Director 2.0 | video | 1 | 0 | 0.0 | No recent timing sample | Single-worker risk |
| LTX-2.3 Audio | video | 1 | 0 | 0.0 | No recent timing sample | Single-worker risk |
Bring your own runtime
The worker runs beside your inference service. It does not replace your runtime, upload your model files, or silently advertise every model it discovers.
Compatibility
Text · Detected automatically
Text · OpenAI-compatible endpoint
Text · OpenAI-compatible endpoint
Text · Detected or entered locally
Text · Detected or entered locally
Text · Operator-entered endpoint
Image / video · Existing ComfyUI bridge; supported models and workflows
Audio · Reviewed direct-runtime profile
Open describes protocol compatibility, not every model or hardware configuration. Check your platform above. ComfyUI uses the existing bridge and supported workflows; the managed installer is separate. Qualification requires a reviewed Grid profile and canary before that capability may be advertised.
Text operators control the advertised model, response limits, concurrency, and operating schedule. Pause without uninstalling.
Model weights stay in your runtime. The worker needs a scoped Grid credential, never a wallet private key.
Community workers process plaintext prompts and outputs. Do not send secrets or regulated data until confidential execution is independently verified and available.
Needed right now
The first managed media profile stays closed until accepted evidence covers every required GPU class. Each evidence set runs three local canaries and cannot enroll a worker or earn rewards.
12 GB to under 20 GB VRAM
0/1 accepted evidence sets
20 GB to under 80 GB VRAM
0/1 accepted evidence sets
80 GB or more VRAM
0/1 accepted evidence sets
Profile ace-step-v1.5-xl-turbo v0.2.4 · status updated Sep 4, 2026 UTC
Public operator proof
Verify that workers are visible and payouts are settling. The Grid does not turn hardware specs or historical demand into guaranteed earnings.
Payout scenario
Public payout evidence is unavailable, so no estimate is shown.
Verify payouts on BaseLive worker check
Paste the exact worker name or worker ID printed after connection. The check reads the public online registry and sends no hardware details.
Start the worker first, then use the identity from its local dashboard or connection log.
Local decision
The manager detects GPU, VRAM, driver, RAM, disk, and architecture locally. The Grid receives a capability tier and measured performance, not your complete hardware inventory.
Exact source, dependencies, model files, and recipes are verified before installation and again before serving. A capability stays unavailable until its local canary passes.
Operator flow
Start the runtime you already use, or install one separately. The worker detects its endpoint and available models locally.
Select what to advertise, set capacity limits, and run a real local generation before the worker connects to the Grid.
Confirm the worker appears online, then serve compatible jobs. Completed accepted work contributes to the current payout split.