Claude Code running with FCC.
macOS/Linux:
curl -fsSL "https://raw.githubusercontent.com/Alishahryar1/free-claude-code/main/scripts/install.sh" | sh
Windows PowerShell:
& ([scriptblock]::Create((irm "https://raw.githubusercontent.com/Alishahryar1/free-claude-code/main/scripts/install.ps1")))
Re-run the same command whenever you want to update. You can review the installers before running them: install.sh and install.ps1.
The installer asks which coding agents to install or verify. Choose at least one; skipped agents are left unchanged. It can also install and configure RTK globally for the selected agents; RTK is off by default.
Open Free Claude Code from your desktop or Start menu.
Open Free Claude Code from your desktop or Applications folder.
Run:
fcc-server
On Windows and macOS, FCC runs in the system tray or menu bar without opening a terminal. Use its menu to open Admin, check server status, restart, or quit. On Windows, left-clicking the tray icon opens Admin directly.
To print the installed Free Claude Code version without starting the server,
run fcc-server --version.
When using fcc-server, keep the terminal open. The Admin UI opens in your
browser after startup by default. Its address is also shown in the log:
INFO: Admin UI: http://127.0.0.1:8082/admin (local-only)
Use the port shown in your terminal if it differs from 8082.
NVIDIA_NIM_API_KEY.MODEL on the default nvidia_nim/nvidia/nemotron-3-super-120b-a12b, or search the model dropdown and select another model.FCC stores Admin settings in ~/.fcc/.env. To require a bearer token on the
proxy API, enable Proxy Authentication in Admin; disabling enforcement keeps
the same token available to FCC launchers.
Claude Code:
fcc-claude
Codex:
fcc-codex
Pi:
fcc-pi
OpenCode:
fcc-opencode
All four launchers use the current Admin UI settings. Use the agent's model picker to choose from the models FCC exposes. Normal CLI arguments still work, for example:
fcc-codex exec "hello"
fcc-pi and fcc-opencode register FCC only for the launched process; your
existing agent settings, sessions, credentials, and extensions remain unchanged.
Select an FCC model from Claude Code's native /model picker.
MODEL dropdown and select a model. If the provider cannot list
models, enter <provider-id>/<exact-provider-model-id> manually.| Provider | Admin UI setting | Example MODEL |
|---|---|---|
| NVIDIA NIM | NVIDIA_NIM_API_KEY | nvidia_nim/nvidia/nemotron-3-super-120b-a12b |
| OpenRouter | OPENROUTER_API_KEY | open_router/openrouter/free |
| Groq | GROQ_API_KEY | groq/llama-3.3-70b-versatile |
| ClinePass | CLINE_API_KEY | cline_pass/cline-pass/kimi-k3 |
| OpenAI / ChatGPT | Connect ChatGPT in the Admin UI | openai/<model-id> |
| xAI (Grok) | XAI_API_KEY | xai/grok-4.5 |
| QwenCloud Token Plan | QWENCLOUD_API_KEY | qwencloud/qwen3.7-plus |
| QwenCloud Coding Plan | QWENCLOUD_CODING_API_KEY | qwencloud_coding/qwen3.7-plus |
| Together AI | TOGETHER_API_KEY | together/zai-org/GLM-5.2 |
| DeepInfra | DEEPINFRA_API_KEY | deepinfra/deepseek-ai/DeepSeek-V4-Flash |
| SiliconFlow | SILICONFLOW_API_KEY | siliconflow/Qwen/Qwen3-32B |
| Nebius Token Factory | NEBIUS_API_KEY | nebius/Qwen/Qwen3-30B-A3B |
| Chutes | CHUTES_API_KEY | chutes/Qwen/Qwen3-32B-TEE |
| Featherless AI | FEATHERLESS_API_KEY | featherless/Qwen/Qwen3-32B |
| Agnes AI | AGNES_API_KEY | agnes/agnes-2.0-flash |
| ZenMux | ZENMUX_API_KEY | zenmux/deepseek/deepseek-v4-flash-free |
| W&B Inference | WANDB_API_KEY | wandb/openai/gpt-oss-20b |
| Azure OpenAI | AZURE_OPENAI_API_KEY and AZURE_OPENAI_BASE_URL | azure_openai/<deployment-name> |
| Google AI Studio (Gemini) | GEMINI_API_KEY | gemini/models/gemini-3.1-flash-lite |
| Google Vertex AI | VERTEX_PROJECT_ID + ADC | vertex/google/gemini-3.5-flash |
| DeepSeek | DEEPSEEK_API_KEY | deepseek/deepseek-chat |
| Mistral La Plateforme | MISTRAL_API_KEY | mistral/devstral-small-latest |
| Mistral Codestral | CODESTRAL_API_KEY | mistral_codestral/codestral-latest |
| OpenCode Zen | OPENCODE_API_KEY | opencode_zen/gpt-5.3-codex |
| OpenCode Go | OPENCODE_API_KEY | opencode_go/minimax-m2.7 |
| Vercel AI Gateway | AI_GATEWAY_API_KEY | vercel/openai/gpt-5.5 |
| Amazon Bedrock | AWS_BEARER_TOKEN_BEDROCK | bedrock/openai.gpt-oss-120b |
| Hugging Face Inference Providers | HUGGINGFACE_API_KEY | huggingface/Qwen/Qwen3-Coder-480B-A35B-Instruct:fastest |
| Cohere | COHERE_API_KEY | cohere/command-a-plus-05-2026 |
| GitHub Models | GITHUB_MODELS_TOKEN | github_models/openai/gpt-4.1 |
| Wafer | WAFER_API_KEY | wafer/DeepSeek-V4-Pro |
| Kimi API | KIMI_API_KEY | kimi/kimi-k2.5 |
| Kimi Code | KIMI_CODE_API_KEY | kimi_code/k3 |
| MiniMax | MINIMAX_API_KEY | minimax/MiniMax-M3 |
| Cerebras Inference | CEREBRAS_API_KEY | cerebras/gpt-oss-120b |
| SambaNova | SAMBANOVA_API_KEY | sambanova/Meta-Llama-3.3-70B-Instruct |
| Kilo.ai | KILO_API_KEY | kilo/kilo-auto/free |
| Fireworks AI | FIREWORKS_API_KEY | fireworks/accounts/fireworks/models/llama-v3p3-70b-instruct |
| Novita AI | NOVITA_API_KEY | novita/deepseek/deepseek-v4-flash-0731 |
| Cloudflare Workers AI | CLOUDFLARE_API_TOKEN and CLOUDFLARE_ACCOUNT_ID | cloudflare/@cf/moonshotai/kimi-k2.6 |
| Z.ai Coding Plan | ZAI_API_KEY | zai/glm-5.2 |
| Z.ai API (pay as you go) | ZAI_API_KEY | zai_api/glm-4.7-flash |
| TokenRouter | TOKENROUTER_API_KEY | tokenrouter/moonshotai/kimi-k3-free |
| NaraRoute | NARAROUTE_API_KEY | nararoute/kimi-k3-free |
| Ollama Cloud | OLLAMA_API_KEY | ollama_cloud/qwen3-coder:480b |
| LM Studio | LM_STUDIO_BASE_URL | lmstudio/<model-id> |
| llama.cpp | LLAMACPP_BASE_URL | llamacpp/<model-id> |
| Ollama | OLLAMA_BASE_URL | ollama/<model-tag> |
AZURE_OPENAI_BASE_URL to its complete v1 endpoint, such as
https://YOUR-RESOURCE-NAME.openai.azure.com/openai/v1/, and select a
deployment that supports Chat Completions. Enter the deployment name as a
custom model slug if it does not appear in the model dropdown.kimi_code/; Kimi API credit keys use
kimi/. Kimi Code plans are for personal interactive coding-agent use under
Kimi's community guidelines.qwencloud_coding/; QwenCloud Token Plan keys
use qwencloud/. The keys and endpoints are not interchangeable. Coding Plan
is for local, personal, interactive coding-agent use under the
Coding Plan terms.OPENCODE_API_KEY but use the explicit
opencode_zen/ and opencode_go/ model prefixes.BEDROCK_BASE_URL to the URL for the same region as
the API key and select one of the listed models.gcloud auth application-default login once; service-account
files and attached service accounts also work. Set VERTEX_PROJECT_ID, and
optionally change VERTEX_LOCATION from its global default.ollama/ prefix.Start LM Studio's local server, load a tool-capable model, and use the model identifier shown by LM Studio with the lmstudio/ prefix. The default URL is http://localhost:1234/v1.
Start llama-server with its OpenAI-compatible Chat Completions API and enough context for the model. Use the local model ID with the llamacpp/ prefix. LLAMACPP_BASE_URL defaults to http://localhost:8080/v1; FCC accepts either the server root or an explicit /v1 suffix.
ollama pull llama3.1
ollama serve
Use the tag shown by ollama list with the ollama/ prefix. OLLAMA_BASE_URL defaults to http://localhost:11434; FCC accepts either the root URL or an explicit /v1 suffix.
MODEL is the fallback for every request. Select a model for MODEL_FABLE, MODEL_OPUS, MODEL_SONNET, or MODEL_HAIKU to override an individual Claude Code tier; select None to use MODEL.
For example, route Opus to nvidia_nim/nvidia/nemotron-3-super-120b-a12b, Sonnet to open_router/openrouter/free, Haiku to lmstudio/qwen3.5-coder, and keep MODEL on zai/glm-5.2.
Open Admin UI → Model Config → Reasoning and select the behavior you want.
| Selection | Behavior |
|---|---|
| From client (default) | Use the effort sent by Claude Code, Codex, Pi, or OpenCode. If none is sent, keep the provider default. |
| Off | Request reasoning to be disabled. |
| Low, Medium, High, X-High, or Max | Override the client with the selected reasoning level. |
| Inherit (Fable, Opus, Sonnet, and Haiku only) | Use the root Reasoning selection. |
Providers that do not support a selected control retain their own behavior.
For terminal use, start fcc-server, then run fcc-claude, fcc-codex,
fcc-pi, or fcc-opencode. Use the guides below for editor integrations.
Install the Claude Code extension. Open VS Code's user settings as JSON and add:
"claudeCode.disableLoginPrompt": true,
"claudeCode.environmentVariables": [
{ "name": "ANTHROPIC_BASE_URL", "value": "http://localhost:8082" },
{ "name": "ANTHROPIC_AUTH_TOKEN", "value": "freecc" },
{ "name": "CLAUDE_CODE_ENABLE_GATEWAY_MODEL_DISCOVERY", "value": "1" },
{ "name": "CLAUDE_CODE_AUTO_COMPACT_WINDOW", "value": "190000" },
{ "name": "DISABLE_AUTOUPDATER", "value": "1" },
{ "name": "DISABLE_FEEDBACK_COMMAND", "value": "1" },
{ "name": "DISABLE_ERROR_REPORTING", "value": "1" }
]
Match the port and authentication token to the Admin UI, then reload the extension.
Start FCC, then edit your Codex configuration:
%USERPROFILE%\.codex\config.toml~/.codex/config.tomlAdd the matching model-catalog path and replace YOUR_USERNAME.
Windows:
model_catalog_json = "C:/Users/YOUR_USERNAME/.fcc/codex-model-catalog.json"
macOS:
model_catalog_json = "/Users/YOUR_USERNAME/.fcc/codex-model-catalog.json"
Then add the shared FCC settings:
model_provider = "fcc"
model = "nvidia_nim/nvidia/nemotron-3-super-120b-a12b"
[model_providers.fcc]
name = "Free Claude Code"
base_url = "http://127.0.0.1:8082/v1"
wire_api = "responses"
[model_providers.fcc.auth]
command = "fcc-codex"
args = ["--print-proxy-auth-token"]
Match the model and port to the Admin UI. The auth command reads FCC's current proxy token automatically. Restart the Codex App after setup or model changes, then select an FCC model from its model picker.
Install the Codex extension. Create or edit ~/.codex/config.toml (%USERPROFILE%\.codex\config.toml on Windows):
model_provider = "fcc"
model = "nvidia_nim/nvidia/nemotron-3-super-120b-a12b"
[model_providers.fcc]
name = "Free Claude Code"
base_url = "http://127.0.0.1:8082/v1"
wire_api = "responses"
[model_providers.fcc.auth]
command = "fcc-codex"
args = ["--print-proxy-auth-token"]
Match model and the port to the Admin UI. The auth command reads FCC's current
proxy token automatically. Restart VS Code after setup or model changes. For
WSL-backed Codex, edit the file inside WSL.
Edit the installed Claude ACP configuration:
C:\Users\%USERNAME%\AppData\Roaming\JetBrains\acp-agents\installed.json~/.jetbrains/acp.jsonSet the environment for acp.registry.claude-acp:
"env": {
"ANTHROPIC_BASE_URL": "http://localhost:8082",
"ANTHROPIC_AUTH_TOKEN": "freecc",
"CLAUDE_CODE_ENABLE_GATEWAY_MODEL_DISCOVERY": "1",
"CLAUDE_CODE_AUTO_COMPACT_WINDOW": "190000",
"DISABLE_AUTOUPDATER": "1",
"DISABLE_FEEDBACK_COMMAND": "1",
"DISABLE_ERROR_REPORTING": "1"
}
Match the port and token to the Admin UI, then restart the IDE.
If Claude Code asks you to log in after you configure the FCC URL and token, open its state file:
%USERPROFILE%\.claude.json~/.claude.jsonMerge this property into the existing JSON without removing its other fields:
"hasCompletedOnboarding": true
If the file does not exist, create it with a complete JSON object:
{
"hasCompletedOnboarding": true
}
Restart Claude Code or the IDE after saving the file.
Configure integrations from Admin UI → Messaging, then click Validate and Apply.
/clear can remove
user prompts.| Usage | Behavior |
|---|---|
/stats | Show session state. |
Standalone /stop | Cancel all work. |
Reply with /stop | Cancel only the selected request while other queued requests continue. |
Standalone /clear | Reset all FCC state and remove every tracked message in that chat, including user prompts, voice notes, FCC replies, Telegram's online notice, and the clear command itself. |
Reply with /clear | Delete the selected message and its literal platform reply subtree while preserving its ancestors and siblings. |
Choose the voice backend you want, then re-run the installer with its option.
| Voice backend | macOS/Linux option | Windows option |
|---|---|---|
| NVIDIA NIM transcription | --voice-nim | -VoiceNim |
| Local Whisper on CPU or CUDA | --voice-local | -VoiceLocal |
| Both backends | --voice-all | -VoiceAll |
| Local Whisper with CUDA 13.0 | --voice-local --torch-backend cu130 | -VoiceLocal -TorchBackend cu130 |
The examples below install NVIDIA NIM transcription. To use another backend, replace the final option with the matching one from the table.
macOS/Linux:
curl -fsSL "https://raw.githubusercontent.com/Alishahryar1/free-claude-code/main/scripts/install.sh" | sh -s -- --voice-nim
Windows PowerShell:
& ([scriptblock]::Create((irm "https://raw.githubusercontent.com/Alishahryar1/free-claude-code/main/scripts/install.ps1"))) -VoiceNim
Restart fcc-server. In Admin UI → Messaging → Voice, enable voice notes, select cpu, cuda, or nvidia_nim, and choose the Whisper model. Local gated models need HUGGINGFACE_API_KEY; NVIDIA NIM transcription needs NVIDIA_NIM_API_KEY.
Re-run the matching command from Install Or Update.
Stop every running FCC command before uninstalling.
Removes
~/.fcc/Keeps
macOS/Linux:
curl -fsSL "https://raw.githubusercontent.com/Alishahryar1/free-claude-code/main/scripts/uninstall.sh" | sh
Windows PowerShell:
& ([scriptblock]::Create((irm "https://raw.githubusercontent.com/Alishahryar1/free-claude-code/main/scripts/uninstall.ps1")))
MIT License. See LICENSE for details.
(top 30 of 41)
Python
96.1%
PowerShell
1.3%
Shell
1.1%
Claude Code running with FCC.
macOS/Linux:
curl -fsSL "https://raw.githubusercontent.com/Alishahryar1/free-claude-code/main/scripts/install.sh" | sh
Windows PowerShell:
& ([scriptblock]::Create((irm "https://raw.githubusercontent.com/Alishahryar1/free-claude-code/main/scripts/install.ps1")))
Re-run the same command whenever you want to update. You can review the installers before running them: install.sh and install.ps1.
The installer asks which coding agents to install or verify. Choose at least one; skipped agents are left unchanged. It can also install and configure RTK globally for the selected agents; RTK is off by default.
Open Free Claude Code from your desktop or Start menu.
Open Free Claude Code from your desktop or Applications folder.
Run:
fcc-server
On Windows and macOS, FCC runs in the system tray or menu bar without opening a terminal. Use its menu to open Admin, check server status, restart, or quit. On Windows, left-clicking the tray icon opens Admin directly.
To print the installed Free Claude Code version without starting the server,
run fcc-server --version.
When using fcc-server, keep the terminal open. The Admin UI opens in your
browser after startup by default. Its address is also shown in the log:
INFO: Admin UI: http://127.0.0.1:8082/admin (local-only)
Use the port shown in your terminal if it differs from 8082.
NVIDIA_NIM_API_KEY.MODEL on the default nvidia_nim/nvidia/nemotron-3-super-120b-a12b, or search the model dropdown and select another model.FCC stores Admin settings in ~/.fcc/.env. To require a bearer token on the
proxy API, enable Proxy Authentication in Admin; disabling enforcement keeps
the same token available to FCC launchers.
Claude Code:
fcc-claude
Codex:
fcc-codex
Pi:
fcc-pi
OpenCode:
fcc-opencode
All four launchers use the current Admin UI settings. Use the agent's model picker to choose from the models FCC exposes. Normal CLI arguments still work, for example:
fcc-codex exec "hello"
fcc-pi and fcc-opencode register FCC only for the launched process; your
existing agent settings, sessions, credentials, and extensions remain unchanged.
Select an FCC model from Claude Code's native /model picker.
MODEL dropdown and select a model. If the provider cannot list
models, enter <provider-id>/<exact-provider-model-id> manually.| Provider | Admin UI setting | Example MODEL |
|---|---|---|
| NVIDIA NIM | NVIDIA_NIM_API_KEY | nvidia_nim/nvidia/nemotron-3-super-120b-a12b |
| OpenRouter | OPENROUTER_API_KEY | open_router/openrouter/free |
| Groq | GROQ_API_KEY | groq/llama-3.3-70b-versatile |
| ClinePass | CLINE_API_KEY | cline_pass/cline-pass/kimi-k3 |
| OpenAI / ChatGPT | Connect ChatGPT in the Admin UI | openai/<model-id> |
| xAI (Grok) | XAI_API_KEY | xai/grok-4.5 |
| QwenCloud Token Plan | QWENCLOUD_API_KEY | qwencloud/qwen3.7-plus |
| QwenCloud Coding Plan | QWENCLOUD_CODING_API_KEY | qwencloud_coding/qwen3.7-plus |
| Together AI | TOGETHER_API_KEY | together/zai-org/GLM-5.2 |
| DeepInfra | DEEPINFRA_API_KEY | deepinfra/deepseek-ai/DeepSeek-V4-Flash |
| SiliconFlow | SILICONFLOW_API_KEY | siliconflow/Qwen/Qwen3-32B |
| Nebius Token Factory | NEBIUS_API_KEY | nebius/Qwen/Qwen3-30B-A3B |
| Chutes | CHUTES_API_KEY | chutes/Qwen/Qwen3-32B-TEE |
| Featherless AI | FEATHERLESS_API_KEY | featherless/Qwen/Qwen3-32B |
| Agnes AI | AGNES_API_KEY | agnes/agnes-2.0-flash |
| ZenMux | ZENMUX_API_KEY | zenmux/deepseek/deepseek-v4-flash-free |
| W&B Inference | WANDB_API_KEY | wandb/openai/gpt-oss-20b |
| Azure OpenAI | AZURE_OPENAI_API_KEY and AZURE_OPENAI_BASE_URL | azure_openai/<deployment-name> |
| Google AI Studio (Gemini) | GEMINI_API_KEY | gemini/models/gemini-3.1-flash-lite |
| Google Vertex AI | VERTEX_PROJECT_ID + ADC | vertex/google/gemini-3.5-flash |
| DeepSeek | DEEPSEEK_API_KEY | deepseek/deepseek-chat |
| Mistral La Plateforme | MISTRAL_API_KEY | mistral/devstral-small-latest |
| Mistral Codestral | CODESTRAL_API_KEY | mistral_codestral/codestral-latest |
| OpenCode Zen | OPENCODE_API_KEY | opencode_zen/gpt-5.3-codex |
| OpenCode Go | OPENCODE_API_KEY | opencode_go/minimax-m2.7 |
| Vercel AI Gateway | AI_GATEWAY_API_KEY | vercel/openai/gpt-5.5 |
| Amazon Bedrock | AWS_BEARER_TOKEN_BEDROCK | bedrock/openai.gpt-oss-120b |
| Hugging Face Inference Providers | HUGGINGFACE_API_KEY | huggingface/Qwen/Qwen3-Coder-480B-A35B-Instruct:fastest |
| Cohere | COHERE_API_KEY | cohere/command-a-plus-05-2026 |
| GitHub Models | GITHUB_MODELS_TOKEN | github_models/openai/gpt-4.1 |
| Wafer | WAFER_API_KEY | wafer/DeepSeek-V4-Pro |
| Kimi API | KIMI_API_KEY | kimi/kimi-k2.5 |
| Kimi Code | KIMI_CODE_API_KEY | kimi_code/k3 |
| MiniMax | MINIMAX_API_KEY | minimax/MiniMax-M3 |
| Cerebras Inference | CEREBRAS_API_KEY | cerebras/gpt-oss-120b |
| SambaNova | SAMBANOVA_API_KEY | sambanova/Meta-Llama-3.3-70B-Instruct |
| Kilo.ai | KILO_API_KEY | kilo/kilo-auto/free |
| Fireworks AI | FIREWORKS_API_KEY | fireworks/accounts/fireworks/models/llama-v3p3-70b-instruct |
| Novita AI | NOVITA_API_KEY | novita/deepseek/deepseek-v4-flash-0731 |
| Cloudflare Workers AI | CLOUDFLARE_API_TOKEN and CLOUDFLARE_ACCOUNT_ID | cloudflare/@cf/moonshotai/kimi-k2.6 |
| Z.ai Coding Plan | ZAI_API_KEY | zai/glm-5.2 |
| Z.ai API (pay as you go) | ZAI_API_KEY | zai_api/glm-4.7-flash |
| TokenRouter | TOKENROUTER_API_KEY | tokenrouter/moonshotai/kimi-k3-free |
| NaraRoute | NARAROUTE_API_KEY | nararoute/kimi-k3-free |
| Ollama Cloud | OLLAMA_API_KEY | ollama_cloud/qwen3-coder:480b |
| LM Studio | LM_STUDIO_BASE_URL | lmstudio/<model-id> |
| llama.cpp | LLAMACPP_BASE_URL | llamacpp/<model-id> |
| Ollama | OLLAMA_BASE_URL | ollama/<model-tag> |
AZURE_OPENAI_BASE_URL to its complete v1 endpoint, such as
https://YOUR-RESOURCE-NAME.openai.azure.com/openai/v1/, and select a
deployment that supports Chat Completions. Enter the deployment name as a
custom model slug if it does not appear in the model dropdown.kimi_code/; Kimi API credit keys use
kimi/. Kimi Code plans are for personal interactive coding-agent use under
Kimi's community guidelines.qwencloud_coding/; QwenCloud Token Plan keys
use qwencloud/. The keys and endpoints are not interchangeable. Coding Plan
is for local, personal, interactive coding-agent use under the
Coding Plan terms.OPENCODE_API_KEY but use the explicit
opencode_zen/ and opencode_go/ model prefixes.BEDROCK_BASE_URL to the URL for the same region as
the API key and select one of the listed models.gcloud auth application-default login once; service-account
files and attached service accounts also work. Set VERTEX_PROJECT_ID, and
optionally change VERTEX_LOCATION from its global default.ollama/ prefix.Start LM Studio's local server, load a tool-capable model, and use the model identifier shown by LM Studio with the lmstudio/ prefix. The default URL is http://localhost:1234/v1.
Start llama-server with its OpenAI-compatible Chat Completions API and enough context for the model. Use the local model ID with the llamacpp/ prefix. LLAMACPP_BASE_URL defaults to http://localhost:8080/v1; FCC accepts either the server root or an explicit /v1 suffix.
ollama pull llama3.1
ollama serve
Use the tag shown by ollama list with the ollama/ prefix. OLLAMA_BASE_URL defaults to http://localhost:11434; FCC accepts either the root URL or an explicit /v1 suffix.
MODEL is the fallback for every request. Select a model for MODEL_FABLE, MODEL_OPUS, MODEL_SONNET, or MODEL_HAIKU to override an individual Claude Code tier; select None to use MODEL.
For example, route Opus to nvidia_nim/nvidia/nemotron-3-super-120b-a12b, Sonnet to open_router/openrouter/free, Haiku to lmstudio/qwen3.5-coder, and keep MODEL on zai/glm-5.2.
Open Admin UI → Model Config → Reasoning and select the behavior you want.
| Selection | Behavior |
|---|---|
| From client (default) | Use the effort sent by Claude Code, Codex, Pi, or OpenCode. If none is sent, keep the provider default. |
| Off | Request reasoning to be disabled. |
| Low, Medium, High, X-High, or Max | Override the client with the selected reasoning level. |
| Inherit (Fable, Opus, Sonnet, and Haiku only) | Use the root Reasoning selection. |
Providers that do not support a selected control retain their own behavior.
For terminal use, start fcc-server, then run fcc-claude, fcc-codex,
fcc-pi, or fcc-opencode. Use the guides below for editor integrations.
Install the Claude Code extension. Open VS Code's user settings as JSON and add:
"claudeCode.disableLoginPrompt": true,
"claudeCode.environmentVariables": [
{ "name": "ANTHROPIC_BASE_URL", "value": "http://localhost:8082" },
{ "name": "ANTHROPIC_AUTH_TOKEN", "value": "freecc" },
{ "name": "CLAUDE_CODE_ENABLE_GATEWAY_MODEL_DISCOVERY", "value": "1" },
{ "name": "CLAUDE_CODE_AUTO_COMPACT_WINDOW", "value": "190000" },
{ "name": "DISABLE_AUTOUPDATER", "value": "1" },
{ "name": "DISABLE_FEEDBACK_COMMAND", "value": "1" },
{ "name": "DISABLE_ERROR_REPORTING", "value": "1" }
]
Match the port and authentication token to the Admin UI, then reload the extension.
Start FCC, then edit your Codex configuration:
%USERPROFILE%\.codex\config.toml~/.codex/config.tomlAdd the matching model-catalog path and replace YOUR_USERNAME.
Windows:
model_catalog_json = "C:/Users/YOUR_USERNAME/.fcc/codex-model-catalog.json"
macOS:
model_catalog_json = "/Users/YOUR_USERNAME/.fcc/codex-model-catalog.json"
Then add the shared FCC settings:
model_provider = "fcc"
model = "nvidia_nim/nvidia/nemotron-3-super-120b-a12b"
[model_providers.fcc]
name = "Free Claude Code"
base_url = "http://127.0.0.1:8082/v1"
wire_api = "responses"
[model_providers.fcc.auth]
command = "fcc-codex"
args = ["--print-proxy-auth-token"]
Match the model and port to the Admin UI. The auth command reads FCC's current proxy token automatically. Restart the Codex App after setup or model changes, then select an FCC model from its model picker.
Install the Codex extension. Create or edit ~/.codex/config.toml (%USERPROFILE%\.codex\config.toml on Windows):
model_provider = "fcc"
model = "nvidia_nim/nvidia/nemotron-3-super-120b-a12b"
[model_providers.fcc]
name = "Free Claude Code"
base_url = "http://127.0.0.1:8082/v1"
wire_api = "responses"
[model_providers.fcc.auth]
command = "fcc-codex"
args = ["--print-proxy-auth-token"]
Match model and the port to the Admin UI. The auth command reads FCC's current
proxy token automatically. Restart VS Code after setup or model changes. For
WSL-backed Codex, edit the file inside WSL.
Edit the installed Claude ACP configuration:
C:\Users\%USERNAME%\AppData\Roaming\JetBrains\acp-agents\installed.json~/.jetbrains/acp.jsonSet the environment for acp.registry.claude-acp:
"env": {
"ANTHROPIC_BASE_URL": "http://localhost:8082",
"ANTHROPIC_AUTH_TOKEN": "freecc",
"CLAUDE_CODE_ENABLE_GATEWAY_MODEL_DISCOVERY": "1",
"CLAUDE_CODE_AUTO_COMPACT_WINDOW": "190000",
"DISABLE_AUTOUPDATER": "1",
"DISABLE_FEEDBACK_COMMAND": "1",
"DISABLE_ERROR_REPORTING": "1"
}
Match the port and token to the Admin UI, then restart the IDE.
If Claude Code asks you to log in after you configure the FCC URL and token, open its state file:
%USERPROFILE%\.claude.json~/.claude.jsonMerge this property into the existing JSON without removing its other fields:
"hasCompletedOnboarding": true
If the file does not exist, create it with a complete JSON object:
{
"hasCompletedOnboarding": true
}
Restart Claude Code or the IDE after saving the file.
Configure integrations from Admin UI → Messaging, then click Validate and Apply.
/clear can remove
user prompts.| Usage | Behavior |
|---|---|
/stats | Show session state. |
Standalone /stop | Cancel all work. |
Reply with /stop | Cancel only the selected request while other queued requests continue. |
Standalone /clear | Reset all FCC state and remove every tracked message in that chat, including user prompts, voice notes, FCC replies, Telegram's online notice, and the clear command itself. |
Reply with /clear | Delete the selected message and its literal platform reply subtree while preserving its ancestors and siblings. |
Choose the voice backend you want, then re-run the installer with its option.
| Voice backend | macOS/Linux option | Windows option |
|---|---|---|
| NVIDIA NIM transcription | --voice-nim | -VoiceNim |
| Local Whisper on CPU or CUDA | --voice-local | -VoiceLocal |
| Both backends | --voice-all | -VoiceAll |
| Local Whisper with CUDA 13.0 | --voice-local --torch-backend cu130 | -VoiceLocal -TorchBackend cu130 |
The examples below install NVIDIA NIM transcription. To use another backend, replace the final option with the matching one from the table.
macOS/Linux:
curl -fsSL "https://raw.githubusercontent.com/Alishahryar1/free-claude-code/main/scripts/install.sh" | sh -s -- --voice-nim
Windows PowerShell:
& ([scriptblock]::Create((irm "https://raw.githubusercontent.com/Alishahryar1/free-claude-code/main/scripts/install.ps1"))) -VoiceNim
Restart fcc-server. In Admin UI → Messaging → Voice, enable voice notes, select cpu, cuda, or nvidia_nim, and choose the Whisper model. Local gated models need HUGGINGFACE_API_KEY; NVIDIA NIM transcription needs NVIDIA_NIM_API_KEY.
Re-run the matching command from Install Or Update.
Stop every running FCC command before uninstalling.
Removes
~/.fcc/Keeps
macOS/Linux:
curl -fsSL "https://raw.githubusercontent.com/Alishahryar1/free-claude-code/main/scripts/uninstall.sh" | sh
Windows PowerShell:
& ([scriptblock]::Create((irm "https://raw.githubusercontent.com/Alishahryar1/free-claude-code/main/scripts/uninstall.ps1")))
MIT License. See LICENSE for details.
(top 30 of 41)
Python
96.1%
PowerShell
1.3%
Shell
1.1%