Code
Hub
Workspaces
Following
Trending
Connect
MCP
copy
Create free account
hub
/
github.com/JustVugg/colibri
/ types & classes
Types & classes
386 in github.com/JustVugg/colibri
⨍
Functions
5,877
◇
Types & classes
386
↓ 46 callers
Class
APIError
c/openai_server.py:59
↓ 29 callers
Class
EvidenceError
c/tools/mw_ur3_raw_adapter.py:83
↓ 21 callers
Class
Engine
c/openai_server.py:2919
↓ 20 callers
Class
FakeProcess
c/tests/test_openai_server.py:631
↓ 11 callers
Class
APIServer
c/openai_server.py:3338
↓ 11 callers
Class
PlannerGeometry
c/family_registry.py:47
↓ 9 callers
Class
StopFilter
Stream text without exposing a full or partial stop sequence.
c/openai_server.py:2518
↓ 7 callers
Class
FakeEngine
c/tests/test_openai_server.py:30
↓ 7 callers
Class
GenerationScheduler
Bounded FIFO admission for the engine's independent KV contexts.
c/openai_server.py:110
↓ 7 callers
Class
MirrorError
c/tools/mirror_plan.py:35
↓ 7 callers
Class
_StubFixture
Two same-named runtime stubs plus one complete fake backend. Everything is generated and compiled inside a private temporary root whose path
c/tests/test_backend_loader.py:158
↓ 6 callers
Class
LayerReplayCache
Memoize exact per-layer replays; all implemented policies are layer-local.
c/tools/residency_sim.py:532
↓ 6 callers
Class
RansRefusal
A named TRUST-VERIFY-REFUSE failure. `code` matches c/rans.h's rans_err_name strings (E_TRUNCATED, E_TABLE_CRC, ...) plus the container-level
c/tools/rans_format.py:51
↓ 5 callers
Class
RegistryError
c/family_registry.py:11
↓ 5 callers
Class
ThinkingStreamSplit
Split GLM's reasoning marker without leaking markers across stream chunks.
c/openai_server.py:2220
↓ 4 callers
Class
stat
c/colibri.c:626
↓ 3 callers
Class
ClientCancelled
c/openai_server.py:71
↓ 3 callers
Class
ClusterRegistry
c/cluster.py:19
↓ 3 callers
Class
Engine
A serve-mode inkling, speaking the protocol in docs/serve_protocol.md.
c/tests/test_inkling_prefix_serve.py:65
↓ 3 callers
Class
FakeTensor
c/tests/test_olmoe_atomic_output.py:38
↓ 3 callers
Class
FamilyConfigError
c/family_registry.py:15
↓ 3 callers
Class
FixtureBuildError
A stub fixture could not be built — distinct from a loader-contract failure. Carrying the compiler command and its complete output means a broken
c/tests/test_backend_loader.py:87
↓ 3 callers
Class
InklingStreamSplit
Strips Inkling's content markers from the visible stream and withholds <|content_thinking|> sections from `content` (they are reasoning, not a
c/openai_server.py:608
↓ 3 callers
Class
LayerSpec
c/tools/residency_sim.py:48
↓ 3 callers
Class
PolicyStats
c/tools/residency_sim.py:56
↓ 3 callers
Class
Serve
One persistent `SERVE=1` engine process.
c/tests/test_deepseek_v4_prefix.py:34
↓ 3 callers
Class
UnknownFamilyError
c/family_registry.py:19
↓ 2 callers
Class
AccessEvent
c/tools/residency_sim.py:36
↓ 2 callers
Class
ClusterServer
c/cluster.py:103
↓ 2 callers
Class
FamilyCapabilities
c/family_registry.py:28
↓ 2 callers
Class
FamilyDescriptor
c/family_registry.py:55
↓ 2 callers
Class
FamilyLimits
c/family_registry.py:36
↓ 2 callers
Class
MalformedRepackTensorError
A repack-kind tensor (REPACK_KINDS) that CANNOT be repacked: wrong dtype for the fp8 path, or no `_scale_inv` sidecar. Named so the refusal is
c/tools/repack_fp8_passthrough.py:341
↓ 2 callers
Class
OutputWriter
Buffers converted tensors and flushes into new model-NNNNN.safetensors shard files once the buffer gets large -- st.h indexes every *.safetensors
c/tools/convert_olmoe_merged.py:71
↓ 2 callers
Class
Plan
c/tools/placement_balance.py:10
↓ 2 callers
Class
ToolSideband
Request-scoped K3 TOOL frames with the same stop policy as DATA.
c/openai_server.py:2582
↓ 2 callers
Class
stat
c/inkling.c:910
↓ 2 callers
Class
stat
c/tests/test_ssd_probe.c:234
↓ 1 callers
Class
Args
c/tests/test_int3_convert.py:46
↓ 1 callers
Class
Args
The subset of `coli convert`'s namespace that cmd_convert reads.
c/tests/test_convert_routing.py:52
↓ 1 callers
Class
BlockingEngine
c/tests/test_openai_server.py:48
↓ 1 callers
Class
BlockingStream
c/tests/test_openai_server.py:591
↓ 1 callers
Class
Classifier
c/tools/k3_repack.py:83
↓ 1 callers
Class
DelayedProfileEngine
c/tests/test_datapoint.py:226
↓ 1 callers
Class
DiagnosticHarness
c/tools/diag_harness.py:280
↓ 1 callers
Class
EngineRunner
c/tools/diag_harness.py:176
↓ 1 callers
Class
FakeEngine
Emits whatever `script` says, so a test can drive tool syntax through the parser.
c/tests/test_anthropic_messages.py:25
↓ 1 callers
Class
InputChangedError
A source shard recorded done in the resume manifest no longer matches its recorded fingerprint (size + mtime_ns): the input changed between ru
c/tools/repack_fp8_passthrough.py:348
↓ 1 callers
Class
JOBOBJECT_EXTENDED_LIMIT_INFORMATION
c/openai_server.py:2884
↓ 1 callers
Class
K3Ref
c/tools/k3_ref.py:109
↓ 1 callers
Class
MEMORYSTATUSEX
c/resource_plan.py:282
↓ 1 callers
Class
OutputDrift
A run produced different tokens than the session's first run. The sweep only offers quality-preserving scheduling knobs, so any output change
c/autotune.py:374
↓ 1 callers
Class
PlannerUnsupportedError
c/family_registry.py:23
↓ 1 callers
Class
QwenEngine
c/tests/test_datapoint.py:255
↓ 1 callers
Class
ResolvedFamily
c/family_registry.py:123
↓ 1 callers
Class
ResumeManifestError
The resume manifest cannot be trusted as-is -- e.g. its entries predate content-validated resume and carry no size/sha256/fingerprint records to
c/tools/repack_fp8_passthrough.py:355
↓ 1 callers
Class
ScaledReplayCache
c/tools/residency_sim.py:606
↓ 1 callers
Class
Serve
One persistent SERVE=1 kimi_k3 process; stderr goes to a file so the pipe can never fill and deadlock the engine.
c/tests/test_kimi_k3_ckpt.py:37
↓ 1 callers
Class
Shards
c/tools/k3_ref.py:30
↓ 1 callers
Class
SourceRepo
c/tools/repair_mtp_int8.py:65
↓ 1 callers
Class
Zero
Returns zeros_like(input) regardless of extra args. Replacing a decoder layer's self_attn/mlp with this makes the layer a clean identity: h =
c/tools/make_qwen36_oracle.py:40
↓ 1 callers
Class
Zero
c/tools/make_qwen36_tiny.py:62
↓ 1 callers
Class
_ChunkEngine
Engine that emits a caller-supplied chunk sequence, to exercise the streaming reasoning splitter across arbitrary chunk boundaries.
c/tests/test_openai_server.py:2051
↓ 1 callers
Class
_ContextExceededEngine
Engine that rejects the prompt before ACCEPT — on_accept is never called.
c/tests/test_openai_server.py:2226
↓ 1 callers
Class
_DeadlineReader
rfile wrapper enforcing a CUMULATIVE deadline on reading one request. SEC: `timeout` below is per socket operation, so it restarts on every byte.
c/openai_server.py:3426
↓ 1 callers
Class
_ExplodingEngine
ACCEPTs the prompt (committing the streaming 200), then dies mid-generation.
c/tests/test_openai_server.py:2264
↓ 1 callers
Class
_F
c/tools/convert_fp8_to_int4.py:648
↓ 1 callers
Class
_MEMSTATUS
c/tools/datapoint.py:102
↓ 1 callers
Class
sigaction
c/colibri.c:7625
↓ 1 callers
Class
stat
c/kimi_k3.c:403
↓ 1 callers
Class
stat
c/tests/test_st_pread.c:54
↓ 1 callers
Class
statfs
c/colibri.c:11342
Class
web/src/ErrorBoundary.tsx:17
Class
APIHandler
c/openai_server.py:3463
Class
Abl
c/abl.h:50
Class
AcceptFrameTest
#597 item 6: the engine's ACCEPT frame gates the HTTP commit, and its invariants.
c/tests/test_openai_server.py:2165
Class
AllocationTest
c/tests/test_residency_sim.py:265
Class
AllowedHostsTest
#597: the DNS-rebinding guard must accept operator-trusted reverse-proxy Host values, while still rejecting everything else by default.
c/tests/test_openai_server.py:1936
Class
AnalysisCacheTest
c/tests/test_analysis_cache.py:41
Class
Args
env_for_engine reads whatever flags the family cares about; every one not set here reads as "not given", which is what argparse yields.
c/tests/test_cli_output.py:392
Interface
AtlasEntry
web/src/Brain.tsx:8
Class
AtomicConverterOutputTest
c/tests/test_converter_atomic_output.py:24
Class
AtomicOlmoeOutputTest
c/tests/test_olmoe_atomic_output.py:46
Class
AuthorityBindingTest
Pin the raw-file authority check independent of manifest tampering: a corpus/comparisons/policy file that is not byte-identical to the frozen
c/tests/test_mw_ur3_raw_adapter.py:573
Class
AuthoritySha256LiteralPinTest
The module's real-authority binding (AUTHORITY_SHA256) is exercised end to end only by RealFrozenCorpusOptInTest above, which skips whenever M
c/tests/test_mw_ur3_raw_adapter.py:1384
Class
AutotuneIntegrationTest
c/tests/test_autotune.py:401
Class
AutotuneUnitTest
c/tests/test_autotune.py:35
Class
Avx2Rows16Tables
c/deepseek_v4.c:16986
Class
BannerModelLineTest
The banner's third line must describe the model that is loaded. It said "GLM-5.2 · 744B MoE · int4 · streaming CPU" for every checkpoint, bec
c/tests/test_cli_output.py:193
Class
BasePolicy
c/tools/residency_sim.py:232
Class
C
c/download_fp8.py:19
Class
CanonicalizerSlotFitTest
U5 fix round, finding 4: _canonicalize_metadata_order's slot-fit guard was a bare assert -- compiled out under python -O, where an oversized h
c/tests/test_fp8_repack.py:1027
Class
CapSentinelShimTest
c/tests/test_openai_server.py:1121
Class
ChatCapForwardingTest
#379/#386 r2 (F9): `coli chat` on a non-glm model spawns openai_server as its local server. An explicit --cap must ride along on that command
c/tests/test_cli_output.py:121
Interface
ChatMessage
web/src/lib/api.ts:3
Class
ChatThinkingTest
c/tests/test_chat_thinking.py:146
Class
CliOutputLanguageTest
c/tests/test_cli_output.py:18
Class
ClientHangupTest
A client that disconnects mid-response must not print a traceback. `coli chat` polls /health on a 2 s timeout while the model loads and drops
c/tests/test_openai_server.py:1509
Class
ClusterRegistryTests
c/tests/test_cluster.py:13
Class
ClusterShardingParityTest
Local CPU must equal cluster-delegated expert sharding, token-exact.
c/tests/test_cluster_sharding.py:104
next →
1–100 of 386, ranked by callers