Skip to content
r/LocalLLaMA · Communities

opengradient open-sourced veil, a local openai-compatible proxy that routes through a TEE.

3090 here, qwen3 32b at q4 for most things, 14b when i need speed. covers maybe 80% of what i do. the other 20% is long reasoning chains and cross-file refactoring where it just isn't close and i end up back on a hosted model, which means the stuff i most wanted to keep local is exactly the stuff that leaves. been look