Code
Hub
Workspaces
Following
Trending
Connect
MCP
copy
Create free account
hub
/
github.com/disler/live-bench
/ functions
Functions
371 in github.com/disler/live-bench
⨍
Functions
371
◇
Types & classes
47
↳
Endpoints
7
↓ 58 callers
Function
_check
(name: str, passed: bool, checks: list[dict])
apps/benchmarks/pi-coding-agent/_tasks.py:52
↓ 58 callers
Function
_check
(name: str, passed: bool, checks: list[dict])
apps/benchmarks/pi-coding-agent-3way/_tasks.py:52
↓ 23 callers
Function
_iso
(dt: datetime)
apps/backend/tests/test_http_ingest.py:27
↓ 21 callers
Function
_run
Run a shell command in cwd, capturing output.
apps/benchmarks/pi-coding-agent/_tasks.py:44
↓ 21 callers
Function
_run
Run a shell command in cwd, capturing output.
apps/benchmarks/pi-coding-agent-3way/_tasks.py:44
↓ 18 callers
Function
_make_point
( group: str, metric: str, value: float, client_ts: datetime, chart_id: str = "default", )
apps/backend/tests/test_store.py:23
↓ 16 callers
Function
set
(state: ConnectionState)
apps/frontend/src/stores/connection.ts:11
↓ 13 callers
Function
_iso
(dt: datetime)
apps/backend/tests/test_sse_stream.py:136
↓ 12 callers
Function
_emit_meta
(server: str, chart: str, group: str, data: dict)
apps/benchmarks/pi-coding-agent/bench.py:50
↓ 9 callers
Function
_err
(code: str, message: str, http_status: int)
apps/backend/src/live_bench_server/http_ingest.py:36
↓ 9 callers
Function
loadFixture
(name: string)
apps/frontend/src/schemas.test.ts:37
↓ 9 callers
Function
makePoint
(overrides: Partial<DataPointStored>)
apps/frontend/src/stores/session.test.ts:109
↓ 9 callers
Method
start
(self)
apps/benchmarks/peak_rss_test.py:44
↓ 8 callers
Function
_parse_events
Read ``stop_after`` SSE frames from an ``httpx`` streaming response. Ignores heartbeat comment lines (``: ping``). Returns a list of dicts with
apps/backend/tests/test_sse_stream.py:142
↓ 8 callers
Function
mountLanding
()
apps/frontend/src/pages/LandingPage.test.ts:88
↓ 7 callers
Function
_ts
(seconds_offset: float = 0.0)
apps/backend/tests/fixtures_export.py:38
↓ 7 callers
Function
emit
(obj: dict)
apps/benchmarks/context-scaling-3way/_mlx_runner.py:27
↓ 7 callers
Function
emit
(obj: dict)
apps/benchmarks/context-scaling/_mlx_runner.py:27
↓ 7 callers
Function
emit
(obj: dict)
apps/benchmarks/qwen36-vs-qwen35-vs-gemma4/_mlx_runner.py:27
↓ 7 callers
Function
emit
(obj: dict)
apps/benchmarks/qwen35-vs-gemma4/_mlx_runner.py:27
↓ 7 callers
Method
get_benchmark
``GET /benchmarks/{id}`` → parsed JSON dict.
apps/cli/src/live_bench_cli/client.py:106
↓ 6 callers
Function
exposed
(wrapper: ReturnType<typeof mountBar>)
apps/frontend/src/components/AnimatedBar.test.ts:40
↓ 6 callers
Method
ingest
``POST /ingest`` → ``IngestResponse``. Drains any pending outbox first (preserving each line's original ``client_ts``), then sends ``
apps/cli/src/live_bench_cli/client.py:120
↓ 6 callers
Function
mountBar
(value: number)
apps/frontend/src/components/AnimatedBar.test.ts:23
↓ 6 callers
Function
open
()
apps/frontend/src/api/stream.ts:64
↓ 6 callers
Function
safeParse
(raw: unknown, picker: (e: ReturnType<typeof parseSseEvent>) => T | null)
apps/frontend/src/api/stream.ts:54
↓ 6 callers
Function
setState
(state: ConnectionState)
apps/frontend/src/api/stream.ts:52
↓ 5 callers
Function
_build_multichart_config
Compact two-chart × two-group BenchmarkConfig for chart-aware tests.
apps/backend/tests/test_store.py:39
↓ 5 callers
Function
_client_with_transport
Build a ``LiveBenchClient`` whose underlying httpx client uses ``transport``.
apps/cli/tests/test_client.py:60
↓ 5 callers
Function
_make_request
()
apps/cli/tests/test_client.py:33
↓ 5 callers
Function
get_logger
Return a configured logger for the given module name.
apps/backend/src/live_bench_server/logging.py:37
↓ 5 callers
Method
ingest
(self, req: IngestRequest)
apps/cli/tests/test_commands.py:53
↓ 5 callers
Function
multiChartSnapshotFrom
(points: DataPointStored[])
apps/frontend/src/stores/session.test.ts:133
↓ 5 callers
Method
publish
Fan out ``event`` to every subscriber for ``session_id``. Synchronous — called from request handlers running inside the asyncio loop.
apps/backend/src/live_bench_server/broker.py:50
↓ 5 callers
Function
rowSessionIds
(wrapper: ReturnType<typeof mount>)
apps/frontend/src/pages/LandingPage.test.ts:107
↓ 5 callers
Function
snapshotFrom
(points: DataPointStored[])
apps/frontend/src/stores/session.test.ts:123
↓ 4 callers
Function
_emit_meta
(server: str, chart: str, group: str, data: dict)
apps/benchmarks/pi-coding-agent-3way/bench.py:54
↓ 4 callers
Function
_seed_multichart
Upsert the multi-chart session into the TestClient's DB.
apps/backend/tests/test_http_ingest.py:75
↓ 4 callers
Method
all_charts
Return every chart in this benchmark as a flat list. For the multi-chart form, returns ``self.charts`` directly (each chart already c
apps/backend/src/live_bench_server/models.py:168
↓ 4 callers
Method
close
(self)
apps/cli/tests/test_commands.py:44
↓ 4 callers
Function
datasetKey
(chart_id: string, group_key: string)
apps/frontend/src/stores/session.ts:58
↓ 4 callers
Method
list_benchmarks
``GET /benchmarks`` → parsed JSON dict.
apps/cli/src/live_bench_cli/client.py:99
↓ 4 callers
Function
load
()
apps/frontend/src/stores/benchmarks.ts:23
↓ 4 callers
Function
wait_stable
Wait until available RAM stabilizes (N consecutive readings within tolerance).
apps/benchmarks/memory_test.py:38
↓ 3 callers
Function
_ok_response
()
apps/cli/tests/test_client.py:50
↓ 3 callers
Method
_parse_error_body
(r: httpx.Response)
apps/cli/src/live_bench_cli/client.py:182
↓ 3 callers
Method
_raise_for_status
(self, r: httpx.Response)
apps/cli/src/live_bench_cli/client.py:191
↓ 3 callers
Function
build_sample_config
Build a compact but complete ``BenchmarkConfig`` for tests.
apps/backend/tests/conftest.py:32
↓ 3 callers
Method
close
(self)
apps/benchmarks/context-scaling/_harness.py:148
↓ 3 callers
Function
create_app
()
apps/backend/src/live_bench_server/main.py:99
↓ 3 callers
Function
jsonOrThrow
(res: Response)
apps/frontend/src/api/rest.ts:18
↓ 3 callers
Function
now_ms
Current wall-clock time as Unix epoch milliseconds (UTC).
apps/backend/src/live_bench_server/db.py:166
↓ 3 callers
Function
parseSseEvent
(raw: unknown)
apps/frontend/src/schemas.ts:227
↓ 3 callers
Method
reset
``POST /benchmarks/{id}/reset`` → parsed JSON dict.
apps/cli/src/live_bench_cli/client.py:113
↓ 3 callers
Function
unload_all
Unload all Ollama models from VRAM.
apps/benchmarks/memory_test.py:53
↓ 2 callers
Function
_better_label
(better: Any)
apps/cli/src/live_bench_cli/commands/show_cmd.py:27
↓ 2 callers
Function
_build_point_from_line
Parse one JSONL line into a ``DataPoint``. Stamps ``client_ts = now_ts()`` immediately if the line does not carry its own. The stamping happe
apps/cli/src/live_bench_cli/commands/emit_cmd.py:47
↓ 2 callers
Function
_flush
Build + POST an ``IngestRequest``; return (points_sent, groups_completed). On transport error the failure is logged and counted, but the stream
apps/cli/src/live_bench_cli/commands/stream_cmd.py:87
↓ 2 callers
Function
_iter_batch_points
Yield one ``DataPoint`` per non-empty line, stamped at read time.
apps/cli/src/live_bench_cli/commands/emit_cmd.py:92
↓ 2 callers
Method
_outbox_path
()
apps/cli/src/live_bench_cli/client.py:199
↓ 2 callers
Method
_post_with_retries
POST ``body`` to ``path`` with exponential backoff on 5xx. Returns parsed JSON on 2xx. Raises ``LiveBenchError`` immediately on 4xx (
apps/cli/src/live_bench_cli/client.py:143
↓ 2 callers
Function
_wait_stable
Wait for available RAM to stabilize, return average.
apps/benchmarks/qwen36-vs-qwen35-vs-gemma4/bench.py:44
↓ 2 callers
Function
_wait_stable
Wait for available RAM to stabilize, return average.
apps/benchmarks/qwen35-vs-gemma4/bench.py:40
↓ 2 callers
Function
allCharts
(config: BenchmarkConfig)
apps/frontend/src/schemas.ts:274
↓ 2 callers
Function
cellKey
(chart_id: string, group_key: string, metric_key: string)
apps/frontend/src/stores/session.ts:54
↓ 2 callers
Function
datetime_to_ms
Convert a datetime (assumed UTC if naive) to Unix-ms integer.
apps/backend/src/live_bench_server/db.py:176
↓ 2 callers
Method
emit_and_track
Emit all 4 chart values + metadata, then check completion.
apps/benchmarks/context-scaling-3way/bench.py:123
↓ 2 callers
Method
emit_and_track
( self, task_key: str, metric_key: str, pi_result: PiResult, score: fl
apps/benchmarks/pi-coding-agent/bench.py:79
↓ 2 callers
Method
emit_and_track
Emit all 4 chart values + metadata, then check completion.
apps/benchmarks/context-scaling/bench.py:113
↓ 2 callers
Method
emit_and_track
Emit all 3 chart values + metadata, then check if this prompt is now complete.
apps/benchmarks/qwen36-vs-qwen35-vs-gemma4/bench.py:131
↓ 2 callers
Method
emit_and_track
( self, task_key: str, metric_key: str, pi_result: PiResult, score: fl
apps/benchmarks/pi-coding-agent-3way/bench.py:83
↓ 2 callers
Method
emit_and_track
Emit all 3 chart values + metadata, then check if this prompt is now complete.
apps/benchmarks/qwen35-vs-gemma4/bench.py:125
↓ 2 callers
Function
emit_memory
Emit a memory measurement to the memory chart.
apps/benchmarks/qwen36-vs-qwen35-vs-gemma4/bench.py:69
↓ 2 callers
Function
emit_memory
Emit a memory measurement to the memory chart.
apps/benchmarks/qwen35-vs-gemma4/bench.py:65
↓ 2 callers
Function
f1_score
Compute set-level F1 between predicted and ground-truth node sets.
apps/benchmarks/context-scaling-3way/bench.py:102
↓ 2 callers
Function
f1_score
Compute set-level F1 between predicted and ground-truth node sets.
apps/benchmarks/context-scaling/bench.py:92
↓ 2 callers
Function
generate_synthetic
Generate a synthetic GraphWalks-style sample with a 'parents' query. Creates a random directed graph, picks a target node that has at least 1
apps/benchmarks/context-scaling-3way/extract_fixtures.py:86
↓ 2 callers
Function
generate_synthetic
Generate a synthetic GraphWalks-style sample with a 'parents' query. Creates a random directed graph, picks a target node that has at least 1
apps/benchmarks/context-scaling/extract_fixtures.py:86
↓ 2 callers
Function
get_rss_kb
(pid)
apps/benchmarks/peak_rss_test.py:21
↓ 2 callers
Function
measure_memory_after
Compute memory delta from baseline. Returns GB consumed.
apps/benchmarks/qwen36-vs-qwen35-vs-gemma4/bench.py:62
↓ 2 callers
Function
measure_memory_after
Compute memory delta from baseline. Returns GB consumed.
apps/benchmarks/qwen35-vs-gemma4/bench.py:58
↓ 2 callers
Function
measure_memory_before
Take a stable baseline reading of available RAM.
apps/benchmarks/qwen36-vs-qwen35-vs-gemma4/bench.py:57
↓ 2 callers
Function
measure_memory_before
Take a stable baseline reading of available RAM.
apps/benchmarks/qwen35-vs-gemma4/bench.py:53
↓ 2 callers
Function
ms_to_datetime
Convert a Unix-ms integer to a timezone-aware UTC datetime.
apps/backend/src/live_bench_server/db.py:171
↓ 2 callers
Function
ollama_chat
Single Ollama /api/chat call with keep_alive=-1. Always returns the same metrics dict shape as MLXRunner.generate(), so the same `run_warm_is
apps/benchmarks/qwen36-vs-qwen35-vs-gemma4/_harness.py:165
↓ 2 callers
Function
ollama_chat
Single Ollama /api/chat call with keep_alive=-1. Always returns the same metrics dict shape as MLXRunner.generate(), so the same `run_warm_is
apps/benchmarks/qwen35-vs-gemma4/_harness.py:165
↓ 2 callers
Function
parse_client_ts
Parse the ISO-8601 TEXT representation we store for ``client_ts``.
apps/backend/src/live_bench_server/db.py:183
↓ 2 callers
Function
parse_final_answer
Extract node IDs from 'Final Answer: [node1, node2, ...]' in the response.
apps/benchmarks/context-scaling-3way/bench.py:92
↓ 2 callers
Function
parse_final_answer
Extract node IDs from 'Final Answer: [node1, node2, ...]' in the response.
apps/benchmarks/context-scaling/bench.py:82
↓ 2 callers
Function
row_to_datapoint_stored
Build a ``DataPointStored`` from a ``datapoints`` row. ``session_id`` may be supplied explicitly when the SELECT did not include that column
apps/backend/src/live_bench_server/db.py:191
↓ 2 callers
Function
run_warm_isolated
Run a single (prompt, settings) measurement with warm-isolated semantics. Flow: 1. One warmup call with `warmup_prompt` — discarded. This
apps/benchmarks/context-scaling-3way/_harness.py:233
↓ 2 callers
Function
run_warm_isolated
Run a single (prompt, settings) measurement with warm-isolated semantics. Flow: 1. One warmup call with `warmup_prompt` — discarded. This
apps/benchmarks/context-scaling/_harness.py:233
↓ 2 callers
Function
run_warm_isolated
Run a single (prompt, settings) measurement with warm-isolated semantics. Flow: 1. One warmup call with `warmup_prompt` — discarded. This
apps/benchmarks/qwen36-vs-qwen35-vs-gemma4/_harness.py:233
↓ 2 callers
Function
run_warm_isolated
Run a single (prompt, settings) measurement with warm-isolated semantics. Flow: 1. One warmup call with `warmup_prompt` — discarded. This
apps/benchmarks/qwen35-vs-gemma4/_harness.py:233
↓ 2 callers
Method
set_prompt
Store the prompt text for a group (called once per prompt).
apps/benchmarks/qwen36-vs-qwen35-vs-gemma4/bench.py:127
↓ 2 callers
Method
set_prompt
Store the prompt text for a group (called once per prompt).
apps/benchmarks/qwen35-vs-gemma4/bench.py:121
↓ 2 callers
Method
stop
(self)
apps/benchmarks/peak_rss_test.py:49
↓ 1 callers
Method
_append_outbox
Append the raw JSON request body as one line to the outbox. ``body`` is the already-serialised ``IngestRequest`` from the caller, so
apps/cli/src/live_bench_cli/client.py:202
↓ 1 callers
Function
_available_gb
Current system available RAM in GB.
apps/benchmarks/qwen36-vs-qwen35-vs-gemma4/bench.py:39
↓ 1 callers
Function
_available_gb
Current system available RAM in GB.
apps/benchmarks/qwen35-vs-gemma4/bench.py:35
next →
1–100 of 371, ranked by callers