repowiserepowise
Sign in
assafelovic/gpt-researcher
OverviewDocsArchitectureKnowledge GraphFilesCode HealthRefactoring

People & History

CommitsContributorsDecisions
ChatPro
Stats
repowiserepowise
ExplorePricingDocs
Sign inIndex repoIndex your repo free
repowiseassafelovic/gpt-researcher
GP
GP

assafelovic / gpt-researcher

pythonmain5d84d2f63.7K linessynced 2h ago

Code health

8.6out of 10Excellent

This codebase scores 8.6 out of 10 on defect risk, which we rate excellent. It also scores maintainability 9.1 and static performance risk 9.9 out of 10. The three are scored separately and never blended into one number. The files you change most are the weak spot: 7 of 614 files are git hotspots, and they average 2.0.

Full health report →
Documentation404pages33 with model-written prose, 371 built from the indexScore validation17/204.66× baselineOf the 20 lowest-health files, 17 were bug-fixed in the last 6 months — 4.66× the 18% baselineDead code27exportsUnused exports that nothing in the graph reaches
Lines of code
63.7K
Files
614
Symbols
1,606
Modules
18
Languages
11

Recent activity

All commits →

Commits

No commits indexed yet.

Decisions

  • Handle List[dict] context from CURATE_SOURCES=True in deep researchproposed · 2h ago
  • Concatenate all pages in PyMuPDFScraperproposed · 2h ago
  • Fetch arXiv via arxiv Client API resultsproposed · 2h ago
  • Fix MCP transport auth by forwarding connection_headersproposed · 2h ago
  • Truncate document content to avoid embedding token limit errorsproposed · 2h ago
  • Coerce none/null/empty for Optional env fieldsproposed · 2h ago

Needs attention

43 open
  • medium severity. ProposedHandle List[dict] context from CURATE_SOURCES=True in deep researchAuto-proposed decision awaiting review
  • medium severity. ProposedConcatenate all pages in PyMuPDFScraperAuto-proposed decision awaiting review
  • medium severity. ProposedFetch arXiv via arxiv Client API resultsAuto-proposed decision awaiting review
  • medium severity. ProposedFix MCP transport auth by forwarding connection_headersAuto-proposed decision awaiting review
  • medium severity. ProposedTruncate document content to avoid embedding token limit errorsAuto-proposed decision awaiting review

Where the risk concentrates

All 7 hotspots →

Ranked by prior bug fixes and change frequency, mined from full git history rather than from the code alone.

FileChurnPrior fixesMaintainersCommits 90d
deep_research.pygpt_researcher/skills99.7th7bug magnet135
researcher.pygpt_researcher/skills99.5th5bug magnet154
costs.pygpt_researcher/utils95.8th253
base.pygpt_researcher/llm_provider/generic93.9th3386
llm.pygpt_researcher/utils92.1th4364

Composition

Open the graph →
python 45%markdown 26%typescript 12%json 10%javascript 3%Other 5%

Explore this codebase

  • Docs404 pages across 18 modules→
  • ChatAsk this codebase a question and get an answer with its sources→
  • Files614 files with per-file docs, health and history→
  • ArchitectureDependency graph, layers, and 23 entry points→
  • Code healthPer-file scores, 842 open findings, coverage and refactoring targets→
  • Knowledge graphEntities, communities, and the paths between them→
  • Change couplingFiles that keep changing together, mined from commit history→
  • CommitsChange-risk ranked history with agent provenance→
  • ContributorsBus factor, per-file maintainers, and the human/agent split→
  • StatsSize class, origin, lifetime churn, rhythm and records→
  • CostsWhat indexing this snapshot cost, by model and by run→