بررسی تطبیقی فهرست ۳٬۰۳۸ قربانی اعلامشده توسط دولت جمهوری اسلامی با ۳٬۰۰۵ قربانی مستندشده از شبکههای اجتماعی
Cross-Referencing the Islamic Republic's 3,038-Name List Against Our 3,005 Victims Documented from Social Media
دولت جمهوری اسلامی در بهمن ۱۴۰۴ فهرستی شامل ۳٬۰۳۸ نام منتشر کرد و مدعی شد اینها تمامی کشتهشدگان رویدادهای دیماه ۱۴۰۴ هستند. این آرشیو نیز بهطور مستقل، از طریق گردآوری خودکار، استخراج با هوش مصنوعی و بازبینی ویراستاری انسانی، تا تاریخ ۱۲ فروردین ۱۴۰۵ مستندات ۳٬۰۰۵ قربانی را از شبکههای اجتماعی گردآوری کرده است. ما این دو مجموعه داده را با تطبیق چندمرحلهای نامها و سپس بازبینی دستی با هم سنجیدیم.
نتایج تطبیق این دو مجموعه داده تفاوتهای چشمگیری را آشکار میکند. از ۳٬۰۳۸ نام فهرست دولت، ۱٬۶۷۱ نفر در مستندات مستقل ما نیز ثبت شدهاند. ۱٬۱۵۵ نام دیگر تنها در فهرست دولت هستند و هیچ نمایه مستقلی از آنها در دادههای شبکههای اجتماعی ما یافت نشده است. اما ۱٬۸۱۴ قربانی که ما از منابع شبکههای اجتماعی مستند کردهایم، در فهرست دولت اصلاً وجود ندارند. در مجموع، دو منبع دستکم ۴٬۸۵۲ قربانی یکتا را ثبت کردهاند. این رقم کف است، نه سقف؛ هم شبکههای اجتماعی و هم دولت همواره در کشتارهای جمعی، بهویژه در شرایط قطع ارتباطات و سرکوب سازمانیافته اطلاعات، کمتر از واقعیت گزارش میدهند.
In February 2026, the Islamic Republic's government published a list of 3,038 people it acknowledged as killed during the January 2026 #IranMassacre. Independently, this archive has documented 3,005 victims from social media evidence as of 1 April 2026, through automated collection, AI-assisted extraction, and human editorial review. We cross-referenced the two datasets using layered name-matching algorithms followed by manual verification.
The results show significant divergence. Of the government's 3,038 names, 1,671 correspond to victims this archive documented independently. Another 1,155 government names have no corresponding record in our data. Most notably, 1,814 victims documented in our archive from social media sources are entirely absent from the government's list. Combined, the two sources establish a minimum of 4,852 unique victims. This figure represents a floor, not a ceiling; both social media coverage and government disclosures are known to undercount in mass atrocity events.
منابع
Sources
۱. فهرست دولت. دفتر ریاست جمهوری اسلامی در ۱۴ بهمن ۱۴۰۴ فایل اکسلی با عنوان «اطلاعیه دفتر رئیس جمهور درباره حوادث دی ماه ۱۴۰۴» منتشر کرد که شامل ۳٬۰۳۸ نام بود. ویراستاران ویکیپدیا این فایل را در ۱۷ بهمن ۱۴۰۴ از نشانی اصلی در president.ir بایگانی کردند (منبع ویکیپدیا). نسخهای بدون تغییر از فایل اصلی در این سایت نگهداری میشود (SHA-256: 5edb85fd). هر ردیف شامل نام، نام پدر و بخشی از شماره شناسایی ملی است.
۲. مدارک جمعآوریشده توسط این سایت. این آرشیو تا تاریخ ۸ فروردین ۱۴۰۵ (نسخه imn-260328-b6f9) مستندات ۳٬۰۰۵ قربانی را از پستهای عمومی شبکههای اجتماعی گردآوری کرده است. دادهها بهصورت خودکار جمعآوری، با هوش مصنوعی استخراج و با بازبینی انسانی نهایی شدهاند. جزئیات روش کار در صفحه درباره ما آمده است.
1. Government list. On 3 February 2026, the Islamic Republic's presidential office published a spreadsheet of 3,038 names titled "Statement of the President's Office Regarding the Events of January 2026." The file was archived by Wikipedia editors on 6 February 2026 from president.ir (see Wikipedia reference). A mirror is hosted on this site (SHA-256: 5edb85fd). Each row contains a name, father's name, and a partial national identification number.
2. The IranMassacre.net archive. 3,005 persons documented from public social media posts as of 28 March 2026 (snapshot imn-260328-b6f9) through automated collection, AI-assisted extraction, and human editorial review. The methodology, limitations, and editorial standards are described on our about page.
یافتهها
Findings
برای این بررسی، ما تکتک نامهای فهرست دولت را با نمایههای مستقل خود تطبیق دادیم (ر.ک. روششناسی). ما این دادهها را بهصورت خودکار از پستهای عمومی شبکههای اجتماعی جمعآوری کرده، با هوش مصنوعی پردازش کرده و با بازبینی انسانی نهایی کردهایم؛ محدودیتهای ذاتی این رویکرد در بخش محدودیتها آمده است. نتایج چهار دستهاند.
We matched every name in the government's list against our independently compiled records (see methodology). Our records are drawn from public social media sources through automated collection, AI-assisted extraction, and human editorial review; inherent limitations of this approach are detailed in the caveats below. The results fall into four categories.
مجموع دو منبع دستکم ۴٬۸۵۲ فرد یکتا را مستند کردهاند. این سایت ۲٬۹۶۱ نمایه جداگانه دارد.
فاصله ۷۷ نفری میان این دو رقم از آنجاست که در فهرست دولت چند رکورد با نام یکسان و نام پدر متفاوت وجود دارد که همگی به یک نمایه در سایت ما اشاره میکنند. اینها افراد واقعی و جداگانهاند، اما چون در شبکههای اجتماعی نام پدر ذکر نمیشود، تفکیک آنها ممکن نیست.
۷۹ مورد دیگر نیز از تطبیق فازی نام به دست آمدهاند که هنوز قطعی نیستند. اگر همه تأیید شوند، شمار تأییدشدهها از ۱٬۶۷۱ به ۱٬۷۵۰ میرسد و شمار «فقط دولت» از ۱٬۱۵۵ به ۱٬۰۷۶ کاهش مییابد. رقم کل ۴٬۸۵۲ تغییر نمیکند، چون موارد نامطمئن تنها میان دستهها جابهجا میشوند.
Across both sources, at least 4,852 unique individuals have been documented. This site hosts 2,961 individual profile pages.
The 77-person gap between these two figures arises where the government's list contains multiple records with the same name but different fathers, all pointing to a single profile on this site. These are real, separate individuals, but social media posts do not include father names, making it impossible to distinguish them.
A further 79 cases were identified through fuzzy name matching and remain unconfirmed. If all were verified, the confirmed count would rise from 1,671 to 1,750, and government-only names would fall from 1,155 to 1,076. The overall total of 4,852 would not change, as uncertain cases shift between categories but remain in the combined count.
روششناسی
Methodology
ما هر یک از ۳٬۰۳۸ نام موجود در فایل اکسل دولت را با پایگاه داده قربانیانی که به صورت مستقل مستند کرده بودیم، مقایسه کردیم. فرایند تطبیق در سه مرحله انجام شد که هر مرحله دامنه وسیعتری از مقایسه را پوشش میداد:
۱. تطبیق دقیق. در این مرحله نامها ابتدا به یک شکل استاندارد فارسی تبدیل شدند. این کار شامل حذف عناوین و القاب و یکسانسازی تفاوتهای املایی بود. سپس نامها مستقیماً با یکدیگر مقایسه شدند. برای نمونه، «جاویدنام ایران آزاد» و «ایران آزاد» پس از این نرمالسازی به یک شکل واحد تبدیل میشوند. این مرحله ۱٬۵۰۲ تطبیق ایجاد کرد.
۲. تطبیق زیررشتهای. در این مرحله نامهایی بررسی شدند که یکی زیرمجموعه دیگری بود و اجزای اصلی نام با هم همخوانی داشت. برای نمونه، «ایران آزاد فردا» در فهرست دولت با «ایران آزاد» در دادههای ما تطبیق داده شد، چون اجزای اصلی نام یکسان بودند. این مرحله ۲۲۲ تطبیق دیگر شناسایی کرد.
۳. تطبیق فازی. نامهایی که در دو مرحله قبل تطبیق نیافته بودند با استفاده از شباهت تقریبی رشتهای مقایسه شدند. از این مقایسه ۸۳۱ نامزد بالقوه شناسایی شد. سپس هر نامزد توسط یک بررسیکننده بهصورت دستی ارزیابی و به یکی از سه دسته تقسیم شد: تطبیق قطعی، تطبیق احتمالی، یا فرد متفاوت.
تمام تطبیقهای خودکار مراحل اول و دوم با نمایههای موجود در پایگاه داده بازبینی و تأیید شدند. ۷۹ مورد همچنان در دسته نامطمئن باقی ماندهاند و بهصورت جداگانه گزارش شدهاند.
پس از اتمام تطبیق، دو اقدام انجام دادیم. نخست، برای افرادی که در هر دو منبع حضور داشتند، اطلاعات ارائهشده توسط دولت (شامل نام، نام پدر و بخشی از کد ملی) را بهعنوان یک منبع تأییدی اضافه به نمایه آنها افزودیم. دوم، برای ۱٬۱۵۵ نامی که فقط در فهرست دولت وجود داشتند، نمایه جدیدی ایجاد کردیم. تنها منبع این نمایههای جدید سند دولتی است و این موضوع بهروشنی در نمایه آنها مشخص شده است.
این گزارش بهصورت دورهای و با مستندسازی قربانیان بیشتر از شبکههای اجتماعی بهروزرسانی میشود. فهرست دولت ثابت باقی میماند؛ اما دادههای ما همچنان در حال رشد هستند.
We compared every name in the government's 3,038-row spreadsheet against our database of independently documented victims. The matching process used three progressively broader layers:
1. Exact match. Names were normalized to a standard Farsi form (removing titles, standardizing spelling variants) and compared directly. For example, "جاویدنام ایران آزاد" and "ایران آزاد" resolve to the same normalized form. This produced 1,502 matches.
2. Substring match. Names where one form is contained within the other at word boundaries. For example, a government entry "ایران آزاد فردا" matching our record "ایران فردا" when the core name components align. This identified 222 additional matches.
3. Fuzzy match. Remaining unmatched names were compared using approximate string similarity, producing 831 candidates. Each candidate was then manually classified by an operator as a match, possible match, or different person.
All automated matches were validated against existing database records. 79 cases remain classified as uncertain and are reported separately.
Following the cross-reference, we took two actions. First, for persons found in both sources, we augmented their records with the government-provided evidence (name, father's name, and partial national ID) as an additional corroborating source. Second, for the 1,155 names found only in the government's list, we created new entries in this archive with the government record as their sole source, clearly marked as such.
This report is periodically updated as we document additional victims from social media sources. The government's list remains fixed; our records grow.
محدودیتها
Caveats
ارقام ارائهشده در این گزارش باید بهعنوان حداقل مستندشده خوانده شوند، نه شمار کامل و نهایی قربانیان.
درمورد دادههای ما. این آرشیو بر پایه پستهایی ساخته شده که در دوران قطع تقریباً کامل اینترنت و پس از آن در شبکههای اجتماعی منتشر شدهاند. خانوادهها و شاهدانی که امکان اتصال به اینترنت نداشتند، تصمیم به انتشار مطلب نگرفتند، یا محتوایشان حذف شده است، در این دادهها حضور ندارند. فرایند گردآوری ما بر استخراج خودکار و پردازش با هوش مصنوعی متکی است. بازبینی ویراستاری انسانی انجام شده، اما هیچ نمایهای بهتنهایی تأیید مستقل نشده است. همپوشانی ۱٬۶۷۱ نام با فهرست خود دولت پشتوانهای برای صحت بخشی از دادههای ماست، اما به معنای تأیید تمام نمایهها نیست.
درمورد دادههای دولت. فهرست جمهوری اسلامی بدون ارائه روششناسی، ذکر منابع یا هرگونه توضیح درباره شیوه گردآوری نامها منتشر شده است. هیچ نهاد مستقلی برای بررسی صحت یا کامل بودن این فهرست دسترسی نداشته است. غیبت ۱٬۸۱۴ قربانی مستندشده ما از فهرست دولت بهخودیخود نشان میدهد که این فهرست همه قربانیان را در بر نمیگیرد.
درمورد رقم کل. تجربه نشان داده که هم شبکههای اجتماعی و هم دولتها در مستندسازی کشتارهای جمعی کمتر از واقعیت گزارش میدهند. این مسئله بهویژه در شرایط قطع ارتباطات و سرکوب عمدی و سازمانیافته اطلاعات شدیدتر است. ۴٬۸۵۲ نفر مستندشده از مجموع دو منبع، کف آن چیزی است که در حال حاضر از شواهد موجود قابل اثبات است.
The figures presented in this report should be read as a documented minimum, not a comprehensive count.
On our data. This archive draws from social media posts that surfaced during and after a near-total internet blackout. Families and witnesses who could not connect, chose not to post, or had their content removed are not represented. Our collection relies on automated extraction and AI-assisted processing; while human editorial review is applied, individual records have not been independently verified. The overlap of 1,671 names with the government's own list provides a measure of corroboration, but does not constitute verification of the remaining records.
On the government's data. The Islamic Republic's list was published without accompanying methodology, sourcing, or explanation of how names were compiled. No independent body has been granted access to verify its completeness. The 1,814 persons documented in our archive who are absent from the government's list suggest the list does not account for all victims.
On the combined total. Both social media documentation and government disclosures are known to undercount in mass atrocity events, particularly under conditions of communications blackout and deliberate suppression of information. The 4,852 persons documented across both sources represent the lower bound of what can currently be established from available evidence.