Hub subclass inventories
83 nodes_resolved_models(), as_clause_lines(), clause_handoff(), enrich_contract_extraction(), flatten_cuad_clause_labels(), flatten_maud_clause_labels(), infer_merger_consideration(), normalize_consideration() (+75 more)
Rebuilt from 2a212e76 (mailroom v0.7.1 + mailroom-dataset v9 corpus, GT-closure revision 46a4d3c2 + ground truth: pared LLM load, 13-node layered-state pipeline, Gmail triage + relations clerk auxiliary flows, review-resolve tray, judge/arbiter lanes, deterministic field scoring, and the Digital-Mailroom monorepo docs alignment).
This build indexes production src/ only (142 files) so the map follows the live architecture: review_resolve.py, posthoc_gt.py, agent_eval.py, specialist suites, honesty-gap metadata, and the 13-node conveyor (human_review pauses with interrupt()). Tests, notebooks, and .opencode/skills are excluded on purpose.
Navigate the architecture map as layers → modules → symbols. The 13-node conveyor strip at the top jumps to intake_node, classify_node, Lane A/B, Boss, catalog, and archive. Comments (rationale nodes) are hidden until you uncheck them. Double-click a symbol to isolate its neighborhood.
ensure_schema()load_config()async_session()load_env()BaseAgentBaseAgentget_managed_prompt()build_graph()setup_logging()inbox_dir()DocumentState--uses→archive_node()INFERREDsrc/graph/state.py → src/graph/build_graph.pyDocumentState--uses→human_review_node()INFERREDsrc/graph/state.py → src/graph/build_graph.pyDocumentState--uses→intake_node()INFERREDsrc/graph/state.py → src/graph/build_graph.pyDocumentState--uses→resume_from_review()INFERREDsrc/graph/state.py → src/graph/build_graph.pyDocumentState--uses→run_pipeline()INFERREDsrc/graph/state.py → src/graph/build_graph.py_resolved_models(), as_clause_lines(), clause_handoff(), enrich_contract_extraction(), flatten_cuad_clause_labels(), flatten_maud_clause_labels(), infer_merger_consideration(), normalize_consideration() (+75 more)
DocumentState, arbiter_node(), boss_escalation_node(), _build_checkpointer(), build_graph(), _build_handoff_context(), _build_specialist_dispatch(), catalog_write_node() (+47 more)
BaseAgent, BossAgent, ComplianceSpecialist, CorporateRecordsSpecialist, CorrespondenceSpecialist, InsuranceClaimsSpecialist, build_structured_schema(), .adjudicate() (+45 more)
configure(), instrument_openai_client(), is_configured(), _apply_taxonomy_settings(), field_is_ambiguous(), get_type_bands(), warm_embedding_model(), _date_pair_days() (+45 more)
InboxHandler, Watcher, _WatcherLock, WatcherLockHeld, lifespan(), run_pipeline(), claim_file(), is_ingestion_paused() (+43 more)
OpsMonitor, analyze_audit_database(), _check_database(), _embed_watcher_running(), get_audit_trail(), get_document_status(), get_matter(), health() (+42 more)
DocumentManifest, PipelineStage, _document_payload_from_manifest(), lookup_document_endpoint(), _move_rejected_to_failed(), _parse_resolve_payload(), _rate_limit_upload(), _require_token() (+40 more)
get_queue(), _ensure_dirs(), accepted_extensions(), archive_dir(), classified_dir(), failed_dir(), get_base_dir(), _get_config() (+40 more)
AuditLogRecord, DocumentRecord, MatterRecord, Base, _write_review_audit_entry(), _touch_heartbeat(), ._query_catalog(), main() (+36 more)
_catalog_by_trace(), completed_filenames(), enrich_sample_row(), expected_fields_for_sample(), expected_fields_meta(), finalize_report(), find_sample_text(), hf_corpus_honesty() (+35 more)
_RotatingFileSink, _langchain_prompt(), prompt_templates(), default_environment(), load_env(), .__call__(), .__init__(), setup_logging() (+33 more)
ComplianceFilingSpecialist, ContractsSpecialist, CorporateRecordsSpecialist, CorrespondenceSpecialist, _SpecialistBase, get_prompt(), .system_prompt(), .__init__() (+30 more)
classify_image(), clean_prediction(), extract_confidence(), extract_reasoning(), extract_runner_up(), _valid_classes(), build_text_messages(), build_vision_messages() (+26 more)
SorterAgent, .review(), build_structured_schema(), format_sorter_subclass_catalogs(), sorter_subclass_catalog(), valid_sorter_subclasses(), _classification_user_message(), _doc_classes_for_prompt() (+25 more)
CorpusUnavailable, Sample, load_cuad_qa(), load_family_rows(), _contracts_from_annotations(), _contracts_from_txt(), _download(), download_all() (+25 more)
ContractsSpecialist, ._configured_max_input_chars(), ._configured_max_tokens(), ._truncate_input(), .__init__(), agent_uses_vision(), _any_specialist_uses_vision(), is_vision_capable() (+24 more)
after_arbiter(), after_boss(), after_classify(), after_extraction(), after_extraction_gated(), after_human_review(), after_judge(), after_retry_classify() (+24 more)
PDFTranscriber, ._call_llm(), ._call_structured(), ._configured_reasoning_effort(), ._skill_appendix(), .system_prompt(), .system_prompt_with_skills(), ._extract_with_vision() (+23 more)
ensure_field_score_configs(), score_and_log_extraction(), _client(), create_trace_score(), deterministic_verdict_label(), emit_in_pipeline_judge_scores(), emit_pipeline_scores(), ensure_score_configs() (+22 more)
client_kwargs(), flush_langfuse(), get_langfuse_client(), get_trace_id(), install_on_dropped(), instrument_openai_client(), observation(), _optional_float() (+22 more)
RunResult, LegalBenchTask, build_parser(), main(), log_run(), _model_name(), print_summary(), run_task() (+21 more)
BaseAgent, .augmented_system_prompt(), ._build_user_content(), ._call_llm(), ._call_structured(), ._call_vision(), ._call_vision_multi(), ._check_deadline() (+21 more)
RunBudgetExceeded, RunDeadlineExceeded, _bounded(), compute_run_metrics(), check_token_budget(), estimate_cost(), get_deadline_seconds(), get_max_total_output_tokens() (+19 more)
AgentTool, ._tool_context(), _memory_dir(), _memory_path(), recent_context(), record_outcome(), stats(), .__init__() (+19 more)
main(), _slug(), _client(), main(), _parse_since(), sync_logs(), _trace_basics(), _trace_stage() (+15 more)
apply_intake(), _extract_text_from_docx(), _extract_text_from_image(), _extract_text_from_pdf(), _file_sha256(), _file_size(), intake_node(), _read_file_text() (+15 more)
archive_node(), _catalog_upsert(), _emit_stage_audit(), _existing_processing_doc_id(), _finalize_aborted(), human_review_node(), _latest_audit_hash(), _normalize_review_decision() (+15 more)
after_report(), align_class(), _as_float(), class_misses_ground_truth(), collect_review_causes(), expected_class(), expected_field_coverage(), format_causes() (+15 more)
flush_braintrust(), flush(), _attach_field_scoring(), diff_report(), filter_real_samples(), _ground_truth_scores(), _ingest_scores(), main() (+15 more)
ProviderConfig, .__init__(), _check_llm_provider(), get_llm(), get_llm_client(), get_llm_model(), instrument_client(), _build_providers() (+14 more)
active_corpus(), adapt_hub_row(), example_for_class(), example_rows(), examples_by_class(), hub_sample(), load_example_pack(), pipeline_corpora() (+14 more)
_build_evaluator_request(), _build_output_definition(), _build_rule_request(), _client(), _current_evaluator_prompt(), _ensure_llm_connection(), _existing_rule_ids(), main() (+14 more)
append_record(), build_record(), default_log_path(), default_sibling_root(), git_snapshot(), _inside(), regenerate(), _run_python() (+13 more)
_denial_reasons(), determination_consistency_is_quality(), honesty_trace_metadata(), insurance_determination_consistent(), insurance_determination_issues(), insurance_expected_set_is_homogeneous(), insurance_gt_is_homogeneous(), _norm_determination() (+13 more)
ArbiterAgent, ImageExtractor, SorterReviewerAgent, .arbitrate(), .system_prompt(), .extract(), ._fallback_extract(), .system_prompt() (+12 more)
AuditLogEntry, archive_document(), _file_sha256(), build_audit_entry(), compute_audit_hash(), compute_audit_hash_v1(), verify_chain(), _all_chains() (+12 more)
_is_retryable_error(), _is_json_mode_400(), _is_retryable(), _retry_after_seconds(), retry_chat_completion(), _retry_config(), retry_sleep_seconds(), _status_code() (+12 more)
cases_for_agent(), evaluate_agent(), load_fixture_cases(), load_local_pack_cases(), load_manifest_cases(), _mean(), _read_text(), score_case() (+11 more)
get_field_types(), all_local_pack_samples(), compliance_local_samples(), corporate_extraction_samples(), _hydrate(), insurance_contrast_samples(), local_pack_status(), _mean() (+11 more)
FakeLangChainLLM, _FakeStructuredRunner, .bind(), .__init__(), .invoke(), ._make_message(), ._run(), .with_structured_output() (+10 more)
ensure_process_tracing(), _escape(), generate_pdf_from_text(), is_real_sample(), _load_manifest(), prepare_samples(), _dim_summary(), judge_one() (+10 more)
validate_extraction(), apply_classification_guard(), apply_extraction_guard(), guard_classification(), guard_extraction(), _has_substantive_content(), _is_valid_confidence(), _valid_subtypes() (+9 more)
LegalBenchAgent, .answer_binary(), .augmented_system_prompt(), .classify_family(), .__init__(), .system_prompt(), .usage(), agent.py (+8 more)
attach_run_scores(), ensure_score_configs_if_enabled(), _environment(), legalbench_trace(), question_observation(), is_enabled(), pipeline_trace(), langfuse_tracing.py (+8 more)
_NoopLangfuse, _NoopSpan, .create_trace_id(), .flush(), .get_current_trace_id(), .set_current_trace_io(), .shutdown(), .start_as_current_observation() (+7 more)
_apply_sqlite_pragmas(), close_db(), _engine_kwargs(), _ensure_models_imported(), get_engine(), get_session(), _get_sessionmaker(), init_db() (+7 more)
_chunk_config(), _extract_compliance(), _extract_contracts(), _extract_corporate_records(), _extract_correspondence(), _extract_insurance_claims(), _instantiate_specialist(), _run_chunked_extraction() (+7 more)
attach_single_doc_extras(), _numeric_extra(), score_and_log_intake(), score_intake_suite(), score_with_suite(), unwrap_suite_result(), suite_scoring.py, Any (+7 more)
main(), _perturb_date(), _perturb_entity_list(), _perturb_free_text(), _perturb_money(), _perturb_name(), _predictions_for(), calibrate_field_scoring.py (+7 more)
_EvalLangChainLLM, is_classify_call(), user_text_from_messages(), ._classify(), ._evidence_classify(), ._extract(), .__init__(), ._run() (+6 more)
WidgetSpec, _client(), _existing_placements(), json_dumps(), main(), _placement_kwargs(), _score_widget(), _spec_to_request() (+6 more)
equivalent_subtypes(), _binary_f1(), _ece(), _mean(), _safe_div(), score_binary(), score_multiclass(), scoring.py (+6 more)
_MockClient, ensure_dirs(), _collect_documents(), _expectation_for(), _load_manifest_expectations(), main(), _mock_get_llm(), validate_pipeline.py (+5 more)
flush_phoenix(), _init_opentelemetry(), _instrument_openai(), instrument_openai_client(), is_configured(), phoenix_enabled(), phoenix_setup.py, Arize Phoenix tracing backend — local, cost-free default for llm-mailroom.… (+5 more)
CompletenessJudge, ._field_list(), .judge_classification(), .judge_completeness(), .judge_extraction_correctness(), .system_prompt(), ._taxonomy_spec(), ._truncate() (+4 more)
ComplianceFilingExtraction, ContractExtraction, CorporateRecordExtraction, CorrespondenceExtraction, InsuranceClaimExtraction, Matter, get_extraction_schema(), documents.py (+4 more)
list_prompts(), PROMPT_TEMPLATES(), docclass_prompts_enabled(), langchain_prompt_version(), managed_prompt_lookup(), langchain_agents/prompts.py, docclass_mode.py, List all available prompt versions. (+4 more)
MockLegalBenchModel, _hash(), .answer_binary(), .classify_family(), .__init__(), .last_usage(), .usage(), legalbench/mock.py (+3 more)
bootstrap_ci(), _clean(), delta_significance(), _resample_means(), bootstrap.py, Any, Random, Bootstrap confidence intervals and small-sample delta testing. Ported verbatim… (+3 more)
_caption_from_text(), _download(), fetch_atticus(), fetch_legalbench(), fetch_pileoflaw(), main(), _stream_pol_records(), fetch_external_samples.py (+3 more)
classes_match(), normalize_class(), score_exact_classification(), check_contract(), pipeline_class(), classification_scoring.py, Any, Classification KPIs after ``merger_agreement`` became a live MAUD class. Dojo… (+2 more)
cutover_agent(), cutover_all(), list_agents(), list_local_models(), load_config(), main(), recommend_cutover_order(), save_config() (+2 more)
SorterAgent, .classify(), .classify_json(), .__init__(), _invoke_sorter(), _LangChainSorterAgent, Mailroom-configured sorter. - Model/budget defaults come from ``taxonomy.yaml``…, Classify a document, optionally with page images attached. Returns ``(doc_type,… (+1 more)
bool_env(), get_env(), load_env(), require_env(), env_utils.py, Load ``braintrust.env`` then ``.env`` into the environment (idempotent).…, Validate that all given environment variables are set and non-empty. Returns…, Get an environment variable with a default fallback. (+1 more)
_append(), _build_versions(), _rules(), _specialist_rules(), prompts_docclass.py, Docclass prompt variants for every mailroom classification-chain role.…, Pure-appended docclass variant: base is a STRICT PREFIX of the result., Derive every variant from the live production template of that role. (+1 more)
load_ground_truth_labels(), load_hf_rows(), _paginate_viewer(), _scan_cap(), _take_rows(), _viewer_rows(), ``max_scan <= 0`` means unlimited (do not use on the 247k Enron set)., Map filename → {expected, expected_subclass} from config=ground_truth. These… (+1 more)
_client(), _cost_models(), _existing_by_name(), main(), _match_pattern(), _prices_match(), sync_models(), sync_models.py (+1 more)
_check_cost_watchdog(), _fetch_openrouter_prices(), _price_for(), _record_langchain_response(), Warn at $0.15, abort the run at $0.20 (cumulative across all samples)., Mirror _wrap_client's usage/cost accounting for a LangChain response., Fetch live OpenRouter pricing (per-token), normalized to $/M tokens. The…
_Choices, _HintedEvalLangChainLLM, .__init__(), ._hint_from_filename(), .__init__(), Evidence classifier with a filename-hint override. The repository's real sample…
_fingerprint(), _normalize_prediction(), _extract_binary(), Any, yes'/'no' normalization for binary answers (lenient)., Deterministic corpus fingerprint for the sampled rows.
family_classification_prompt_v1(), get_prompt(), legalbench/prompts.py, Versioned LegalBench task prompts. Prompt version = experiment identity in the…, Fill the 25-family list into the multiclass prompt (called per run so the…, Resolve a prompt version to its system-prompt text.
chat, completions, _fake_client(), _fake_judge_client(), .create()
._build_multimodal(), ._uses_vision(), True when this agent's model accepts image input and (optionally) page images…, Build the user-message content for a document input. Vision-capable models get…
_prompt_versions(), _bound_prompt_versions(), Prompt versions bound during the run (best-effort; Langfuse-managed prompts…, Version keys currently wired into production / agent defaults. Used for catalog…
__init__.py package markers are expected.High betweenness centrality (0.088) - this node is a cross-community bridge.
High betweenness centrality (0.051) - this node is a cross-community bridge.
High betweenness centrality (0.046) - this node is a cross-community bridge.
`BaseAgent` has 9 INFERRED edges - model-reasoned connections that need verification.
`BaseAgent` has 10 INFERRED edges - model-reasoned connections that need verification.
`build_graph()` has 25 INFERRED edges - model-reasoned connections that need verification.
1 weakly-connected nodes found - possible documentation gaps or missing edges.