
O1
Gold VerifiedPublisher: EverestQ Engineering · Category: Coding
Parameters
34B
Context Window
128K
Default Size
22.0 GB
Downloads
174,300
High-throughput code generation, automated refactoring, and architectural design engine for complex software projects.
PULL MODEL
eq pull everestq/o1
RUN INFERENCE
eq run everestq/o1 "Write a Go distributed raft consensus engine"
Attention Architecture
Utilizes Grouped-Query Attention (GQA) with 16-token fixed Paged KV Cache block allocation for zero-fragmentation memory residency.
ONNX & EQC Compatible
Compiles directly through the EQC toolchain into standalone .eqx binary packages with fused SwiGLU kernels.
Multi-Hardware Support
Auto-detects CUDA RTX/A100/H100, Apple Silicon Metal Performance Shaders, or AVX-512 CPU execution backends.