در دست ساختIn developmentآغاز ۱۴۰۴Begun 2025

شهربینUrbanus

گراف دانش چندصدایی تهران

A pluralistic knowledge graph of Tehran

این پروژه می‌کوشد ادعاهایی را که پژوهش‌های شهری فارسی و انگلیسی دربارهٔ تهران مطرح کرده‌اند به گرافی دانشی تبدیل کند که هر ادعا را همراه با عبارت اصلی منبع، نقش آن عبارت در مقاله و سابقهٔ داوری‌اش نگه می‌دارد. در این گراف اختلاف‌نظرها حذف نمی‌شوند؛ هر جا منابع یا چارچوب‌های تفسیری با یکدیگر هم‌سو نیستند، هر دو موضع در کنار هم ثبت می‌شود. هیچ خروجی مدل مستقیم وارد گراف نمی‌شود: هر ادعای پیشنهادی در گامی جداگانه با عبارت شاهدش سنجیده می‌شود و باید از قاعده‌هایی صریح بگذرد که در کد بررسی می‌شوند.

Urbanus turns what Persian and English urban research claims about Tehran into a knowledge graph in which every claim keeps its verbatim source passage, the role that passage plays in its paper, and its review history. Disagreement is not resolved away: where sources or interpretive lenses conflict, both positions are recorded side by side. No model output enters the graph directly: each proposed claim is judged against its evidence passage in a separate step and must pass explicit rules that are checked in code.

  • ۵۲۳523مقالهٔ فارسی و انگلیسیpapers
  • ۸۳۳833ادعای داوری‌شدهgoverned claims
  • ۶۵65محله در ۲۲ منطقهneighbourhoods
  • ۲۳23پرسش سنجشcompetency questions
نمای گراف دانش تهران با ۲۲ منطقه، خطوط مترو، لایه‌های تفسیری و ادعاهای مناقشه‌برانگیز / The Tehran knowledge graph with 22 districts, metro lines, interpretive layers and contested claims
نمای گراف در مرداد ۱۴۰۵، هنگامی که ۳۴۵ ادعا داشتThe graph in August 2026, at 345 claims

ایده و تصمیم‌هاIdea and decisions

ایده و چارچوب نظری، یعنی نگه‌داشتن خوانش‌های رقیب در کنار هم، از من است. دامنهٔ پیکره، مفاهیم مناقشه‌برانگیز و قاعده‌های داوری را من نوشتم، هر مرحله را طراحی و هدایت کردم، دستورکارها را تأیید کردم و موارد مبهم را خودم تصمیم گرفتم.

The idea and its framework, keeping rival readings side by side, are mine. I wrote the scope of the corpus, the contested concepts and the rules of review, designed and directed each stage, approved the briefs and decided the open cases myself.

ابزار و روشTools and method

برای ساختن گرافی در این مقیاس در زمانی محدود، مدل‌های زبانی را به‌عنوان ابزار به کار گرفتم، نخست یک مدل محلی و سپس Claude: کدها، استخراج ادعاها از ۵۲۳ مقاله، سنجش هر ادعا با عبارت شاهدش و بازبینی‌های نمونه‌ای با آن‌ها انجام شد. هیچ خروجی مدل بی‌گذر از قاعده‌هایی که در کد بررسی می‌شوند وارد گراف نمی‌شود، و جایی که دانش فنی من کم بود ابزار آن را پر کرد.

To build a graph at this scale in limited time I used language models as the tool, a local model first and then Claude: the code, the extraction of claims from 523 papers, the check of each claim against its evidence passage and the audit passes were done with them. No model output enters the graph without passing rules checked in code, and where my technical knowledge ran short the tool filled it.

  1. گردآوری پیکرهCorpus

    مقاله‌های فارسی و انگلیسی دربارهٔ شهرسازی تهران از پایگاه‌های دسترسی آزاد گردآوری شد.

    Persian and English papers on Tehran’s urbanism, gathered from open-access databases.

  2. غربالگریScreening

    هر مقاله از نظر موضوع، مکان و مقیاس فضایی سنجیده شد تا تنها پژوهش‌های شهری مربوط به تهران باقی بماند.

    Each paper screened for subject, place and spatial scale.

  3. استخراج ادعاExtracting claims

    یافته‌ها از متن کامل مقاله‌ها استخراج شد و هر نقل‌قول به‌طور خودکار با متن اصلی تطبیق داده شد.

    Findings extracted from full texts; every quote checked against its source.

  4. داوریReview

    هر ادعا بر پایهٔ عبارت شاهد خود پذیرفته، اصلاح یا رد شد.

    Each claim accepted, revised or rejected against its evidence passage.

  5. ثبت مناقشهHolding disagreement

    مفاهیم مناقشه‌برانگیزی چون بافت فرسوده و فروش تراکم با هر دو خوانش ثبت می‌شوند.

    Contested concepts, such as worn-out fabric and density selling, kept with both readings.

  6. پیوند مکانیSpatial linking

    ادعاها به فرهنگی جغرافیایی از محله‌ها و مناطق ۲۲گانهٔ تهران پیوند داده شدند.

    Claims linked to a gazetteer of Tehran’s neighbourhoods and 22 districts.

  7. آزمونTesting

    مجموعه‌ای از پرسش‌های سنجش، درستی گراف را پس از هر تغییر می‌آزماید.

    A suite of competency questions tests the graph after every change.

وضعیت: تاکنون ۸۳۳ ادعا از ۲۱۰ مقالهٔ سطح محله و منطقه وارد گراف شده است و استخراج ۲۴۰ مقالهٔ سطح شهر گام بعدی است.Status: 833 claims from 210 neighbourhood and district papers are in the graph; city-level extraction of 240 papers comes next. Tools: OpenAlex, LinkML, RDF and SPARQL, Python.