MCPcopy Create free account

hub / github.com/NJU-LINK/DR3-Eval / functions

Functions786 in github.com/NJU-LINK/DR3-Eval

↓ 1 callersFunctionpreview_tasks
Preview tasks without running them.
src/runners/batch_tasks.py:1312
↓ 1 callersFunctionprint_evaluation_report
Print a human-readable evaluation report.
benchmarks/evaluators/eval_intermediate_insights.py:420
↓ 1 callersFunctionprint_results
Print results
benchmarks/evaluators/calculate_average_score.py:102
↓ 1 callersMethodprint_summary
Print copying summary to console
benchmarks/subset_extraction/gaia-to-text-103-mover.py:130
↓ 1 callersMethodprint_summary
Print grading summary
benchmarks/subset_extraction/gaia-text-103-grader.py:244
↓ 1 callersFunctionprocess_input
Process user input, especially files. Returns formatted initial user message content list and updated task description.
src/io/input_handler.py:63
↓ 1 callersMethodprocess_llm_response
Process LLM response
src/llm/providers/openai_client.py:337
↓ 1 callersFunctionquestion_scorer
( model_answer: str, ground_truth: str, )
benchmarks/evaluators/eval_utils.py:266
↓ 1 callersMethodquick_complete
Quick completion for simple tasks like compression. Uses minimal parameters for fast response. Args: pro
src/llm/base_client.py:336
↓ 1 callersFunctionrender
(ts)
libs/miroflow-tools/src/miroflow_tools/mcp_servers/utils/url_unquote.py:63
↓ 1 callersMethodrerank
Rerank chunks based on relevance to the query. Only filters out irrelevant chunks (below threshold), keeps all relevant ones
libs/miroflow-tools/src/miroflow_tools/tools/rag_rerank.py:96
↓ 1 callersFunctionrestore_folder_db_files
Restore all hidden db files in a folder. NOTE: This function is kept for backward compatibility but is no longer needed since prepar
src/runners/batch_tasks.py:214
↓ 1 callersFunctionrun_batch_tasks
Run multiple tasks in sequence. Args: data_dir: Directory containing task data tasks: List of task dictionaries resul
src/runners/batch_tasks.py:1216
↓ 1 callersFunctionrun_benchmark
Main entry point for running benchmarks with Hydra.
benchmarks/common_benchmark.py:772
↓ 1 callersMethodrun_evaluation
Run the full benchmark evaluation process
benchmarks/common_benchmark.py:722
↓ 1 callersMethodrun_main_agent
Execute the main end-to-end task
src/core/orchestrator.py:816
↓ 1 callersMethodrun_parallel_inference
Run inference on multiple tasks in parallel using threading
benchmarks/common_benchmark.py:413
↓ 1 callersFunctionrun_single_task
Run a single task and save results to a unique run folder. Args: data_dir: Directory containing task data task: Task dictiona
src/runners/batch_tasks.py:878
↓ 1 callersMethodrun_single_task
Run inference for a single benchmark task with pass@k support Args: task: BenchmarkTask object Returns:
benchmarks/common_benchmark.py:119
↓ 1 callersFunctionsafe_unquote
(s: str, encoding="utf-8", errors="ignore")
libs/miroflow-tools/src/miroflow_tools/mcp_servers/utils/url_unquote.py:24
↓ 1 callersFunctionsave_batch_summary
Save batch execution summary to a file in the result folder. This file contains: - Execution statistics - List of tasks that nee
src/runners/batch_tasks.py:1091
↓ 1 callersFunctionsave_final_report
Save the final report (after validation) to the run folder.
src/runners/batch_tasks.py:865
↓ 1 callersFunctionsave_initial_report
Save the initial report (before validation) to the run folder.
src/runners/batch_tasks.py:852
↓ 1 callersFunctionsave_rag_queries_summary
Save RAG queries summary to a human-readable txt file. This file shows all RAG queries organized by agent and turn, making it easy t
src/runners/batch_tasks.py:754
↓ 1 callersMethodsave_results
Save evaluation results to JSONL file
benchmarks/common_benchmark.py:485
↓ 1 callersMethodserialize_for_json
Convert objects to JSON-serializable format
src/logging/task_logger.py:278
↓ 1 callersFunctionset_allowed_db_path
Set the allowed database path for RAG operations. When set, only this path (or paths that resolve to it) can be used for RAG operations.
libs/miroflow-tools/src/miroflow_tools/mcp_servers/rag_mcp_server.py:57
↓ 1 callersFunctionset_retrieval_log_path
Set the path for saving retrieval results.
libs/miroflow-tools/src/miroflow_tools/mcp_servers/rag_mcp_server.py:117
↓ 1 callersFunctionshould_keep_file
(f)
src/io/folder_processor.py:480
↓ 1 callersFunctionsl
(c)
docs/static/js/fontawesome.all.min.js:5
↓ 1 callersFunctionsortTable
* Sort the rows in a HTML Table * * @param Table The Table DOM object * @param col The zero-based column number by which to sort * @param dir
docs/static/js/sort-table.js:21
↓ 1 callersMethodstart_sub_agent_session
Start a new sub-agent session
src/logging/task_logger.py:186
↓ 1 callersMethodsummarize_chunks
Summarize each chunk to extract key information. Args: query: The original search query chunks: List
libs/miroflow-tools/src/miroflow_tools/tools/rag_rerank.py:266
↓ 1 callersFunctiontest_persistent_session
()
libs/miroflow-tools/src/miroflow_tools/mcp_servers/browser_session.py:75
↓ 1 callersMethodto_dict
(self)
evaluators/utils/run_all.py:50
↓ 1 callersFunctiontqdm
(iterable, **kwargs)
eval.py:67
↓ 1 callersMethodupdate_original_files
Update original task files with grading results
benchmarks/subset_extraction/gaia-text-103-grader.py:215
↓ 1 callersFunctionverify_answer_browsecomp
Use BrowseComp judge to verify if the predicted answer is correct.
benchmarks/evaluators/eval_utils.py:402
↓ 1 callersFunctionverify_answer_for_datasets
Verify the answer for a given dataset. Returns a tuple of (result, judge_type).
benchmarks/evaluators/eval_utils.py:513
↓ 1 callersFunctionverify_answer_hle
Use HLE-style LLM judge to verify if the predicted answer is correct. Returns the evaluation result as a string: "CORRECT", "INCORRECT", or "
benchmarks/evaluators/eval_utils.py:172
↓ 1 callersFunctionverify_answer_simpleqa
Use LLM to verify if the predicted answer is correct. Expects the LLM to choose between A (correct), B or C (incorrect).
benchmarks/evaluators/eval_utils.py:113
↓ 1 callersFunctionverify_answer_xbench_deepresearch
Use XBench-DeepResearch judge to verify if the predicted answer is correct.
benchmarks/evaluators/eval_utils.py:461
↓ 1 callersFunctionxc
(c)
docs/static/js/fontawesome.all.min.js:5
↓ 1 callersFunctionyl
(a)
docs/static/js/fontawesome.all.min.js:5
Function$c
()
docs/static/js/fontawesome.all.min.js:5
FunctionAutoplay
(slider)
docs/static/js/bulma-carousel.js:1042
FunctionBreakpoints
(slider)
docs/static/js/bulma-carousel.js:1216
FunctionCoordinate
()
docs/static/js/bulma-carousel.js:329
FunctionEventEmitter
()
docs/static/js/bulma-slider.js:286
FunctionEventEmitter
()
docs/static/js/bulma-carousel.js:280
FunctionFade
(transitioner, slider)
docs/static/js/bulma-carousel.js:2054
FunctionInfinite
(slider)
docs/static/js/bulma-carousel.js:1354
FunctionJ
()
docs/static/js/fontawesome.all.min.js:5
FunctionLoop
(slider)
docs/static/js/bulma-carousel.js:1434
FunctionMp3Converter
(local_path: str, extension: str, **kwargs)
src/io/input_handler.py:902
FunctionNavigation
(slider)
docs/static/js/bulma-carousel.js:1502
FunctionPagination
(slider)
docs/static/js/bulma-carousel.js:1653
FunctionSl
(c)
docs/static/js/fontawesome.all.min.js:5
FunctionSwipe
(slider)
docs/static/js/bulma-carousel.js:1812
FunctionTransitioner
(slider)
docs/static/js/bulma-carousel.js:1954
FunctionTranslate
(transitioner, slider)
docs/static/js/bulma-carousel.js:2190
FunctionVc
(c)
docs/static/js/fontawesome.all.min.js:5
FunctionWavConverter
Converts WAV files to markdown via extraction of metadata (if `exiftool` is installed), and speech transcription (if `speech_recognition` is
src/io/input_handler.py:851
FunctionWikipediaConverter
(local_path: str, **kwargs)
src/io/input_handler.py:444
FunctionYouTubeConverter
(local_path: str, url: str)
src/io/input_handler.py:483
Method__init__
(self)
eval.py:73
Method__init__
(self, error_msg: str)
src/utils/wrapper_utils.py:19
Method__init__
(self, response: Any, extra_info: dict = None)
src/utils/wrapper_utils.py:31
Method__init__
(self, message: str, original_error: Exception = None, trace_id: str = None)
src/llm/exceptions.py:100
Method__init__
(self, content)
src/llm/providers/openai_client.py:623
Method__init__
(self, message, finish_reason)
src/llm/providers/openai_client.py:629
Method__init__
(self, choices, usage)
src/llm/providers/openai_client.py:634
Method__init__
Initialize Progressive Memory. Args: max_tokens: Maximum tokens allowed in memory reserved_tokens: T
src/core/progressive_memory.py:85
Method__init__
Initialize compressor with LLM client. Args: llm_client: LLM client for compression
src/core/progressive_memory.py:354
Method__init__
( self, main_agent_tool_manager: ToolManager, sub_agent_tool_managers: Dict[str, ToolM
src/core/orchestrator.py:66
Method__init__
(self, **options: Any)
src/io/input_handler.py:319
Method__init__
(self, title: Union[str, None] = None, text_content: str = "")
src/io/input_handler.py:400
Method__init__
Initialize ToolManager. :param server_configs: List returned by create_server_parameters()
libs/miroflow-tools/src/miroflow_tools/manager.py:60
Method__init__
Initialize from environment variables
libs/miroflow-tools/src/miroflow_tools/tools/enhanced_audio.py:44
Method__init__
Initialize Enhanced VQA client from environment variables.
libs/miroflow-tools/src/miroflow_tools/tools/enhanced_vqa.py:48
Method__init__
Initialize the reranker. Args: api_key: OpenAI API key (default: from OPENAI_API_KEY env var) base_u
libs/miroflow-tools/src/miroflow_tools/tools/rag_rerank.py:70
Method__init__
Initialize the summarizer. Args: api_key: OpenAI API key (default: from OPENAI_API_KEY env var) base
libs/miroflow-tools/src/miroflow_tools/tools/rag_rerank.py:240
Method__init__
Initialize the post-processor. Args: api_key: OpenAI API key base_url: OpenAI API base URL
libs/miroflow-tools/src/miroflow_tools/tools/rag_rerank.py:415
Method__init__
Initialize RAG Tool. Args: api_key: OpenAI API key base_url: OpenAI API base URL embeddi
libs/miroflow-tools/src/miroflow_tools/tools/rag_tool.py:93
Method__init__
Initialize from environment variables
libs/miroflow-tools/src/miroflow_tools/tools/enhanced_video.py:41
Method__init__
(self, server_params)
libs/miroflow-tools/src/miroflow_tools/mcp_servers/browser_session.py:30
Method__init__
(self, config: Optional[EvalConfig] = None)
evaluators/information_recall.py:51
Method__init__
(self, config: Optional[EvalConfig] = None)
evaluators/format_compliance.py:49
Method__init__
(self, config: Optional[EvalConfig] = None)
evaluators/depth_quality.py:41
Method__init__
(self, config: Optional[EvalConfig] = None)
evaluators/factual_accuracy.py:598
Method__init__
(self, config: Optional[EvalConfig] = None)
evaluators/utils/llm_client.py:241
Method__init__
(self)
evaluators/utils/document_loader.py:155
Method__init__
(self, message: str, metric_name: str = "", details: Dict[str, Any] = None)
evaluators/utils/base.py:27
Method__init__
Initialize the evaluation runner with config and metrics.
evaluators/utils/run_all.py:74
Method__init__
Initialize benchmark evaluator Args: data_dir: Path to benchmark data directory benchmark_name: Name of the
benchmarks/common_benchmark.py:79
Method__init__
Initialize generic evaluator Args: data_dir: Path to benchmark data directory benchmark_name: Name of the be
benchmarks/common_benchmark.py:569
Method__init__
(self, target_path: str, task_per_run: int, data_path: str)
benchmarks/check_progress/common.py:739
Method__init__
Initialize the copier Args: gaia_text_103_data_path: Path to GAIA-Text-103 standardized data file output_dir
benchmarks/subset_extraction/gaia-to-text-103-mover.py:35
Method__init__
Initialize the grader Args: extraction_dir: Directory containing extracted GAIA-Text-103 tasks
benchmarks/subset_extraction/gaia-text-103-grader.py:61
Method__len__
(self)
src/core/progressive_memory.py:339
← previousnext →501–600 of 786, ranked by callers