The AI Pioneers - October 2026
مفارقة أساس النموذج: بناة محركات اكتشاف الذكاء الاصطناعي يديرون مواقع ويب الشركات التي لا تستطيع أنظمة الذكاء الاصطناعي قراءتها بالكاد.
هل الشركات التي تبني الذكاء الاصطناعي الرائد تجعل مواقعها الإلكترونية جاهزة لبحث الذكاء الاصطناعي؟
الخمس شركات في هذه الطبعة تنفق بكثرة لتغيير طريقة العثور على المعلومات. تقوم نماذجها ومساعديها ومنتجات البحث بتعليم المشترين طرح سؤال على آلة بدلاً من تصفح موقع ويب. لقد سألنا سؤالاً بسيطًا: كيف تستعد هذه الشركات لمواقعها الإلكترونية الخاصة لهذا التحول؟
الإجابة هي مفارقة. على المجالات الثلاثة التي يمكننا قياسها، تسمح كل المواقع لزاحفي الذكاء الاصطناعي بالدخول، ومع ذلك لا ينشر أي منها ملف سياق llms.txt ولا يكشف عن هوية شركة منظمة يمكن لمحرك الإجابة الاعتماد عليها. فهي مفتوحة ولكن غير مفسرة. يمكن لنظام الذكاء الاصطناعي الزيارة ولكنه يجب أن يخمن ما يراه.
التناقض داخل جوجل هو الإشارة الأكثر وضوحًا. يحصل مختبر البحث الخاص بها، DeepMind، على درجة 67 في نفس الفحوصات تمامًا، أكثر من ثلاثة أضعاف درجة google.com. يتم بناء الجاهزية داخل فرق المنتجات والبحث، وليس على مستوى الشركة حيث يبدأ المستثمرون والصحفيون والهيئات التنظيمية والمشترون المؤسسيون بحثهم.
لم يتمكن رائدان من الحصول على درجات. قام موقع OpenAI بإعادة توجيه وتحدي الماسح الضوئي الخاص بنا، ولا تنشر شركة SpaceX أي سياسة للمتصفحات على الإطلاق. نحن نبلغ عن هذه النتائج كاستنتاجات وليس كأصفار: عندما لا يمكن التحقق بشكل مستقل من الموقع الإلكتروني للشركة نفسها، تواجه أنظمة الذكاء الاصطناعي نفس الحد، وتعود إلى مصادر ثانوية لا تتحكم فيها.
Synthesis
لم ينشر أي من الرائدين المقاسين ملف llms.txt على نطاقهم الرئيسي.
المواقع الثلاثة المقاسة تسمح لزاحفات الذكاء الاصطناعي ولكن لا تعرض هوية الشركة المنظمة.
حصلت شركة Google DeepMind على 67 في نفس الاختبارات، أي أكثر من ثلاثة أضعاف النطاق الرئيسي.
Directives
Separate access from understanding: Allowing AI crawlers is only the first step. Leadership should ask whether AI systems can also tell who the company is, what it offers and which pages are authoritative. Today, none of the measured pioneers answer that question on their main domain.
Bring sub-brand practice up to corporate level: Where a research lab or product team has already done the work, as DeepMind has, the same standard should apply to the corporate domain. That is the page AI systems cite when buyers, journalists and regulators ask about the company.
Make protection deliberate, not accidental: Bot protection and missing crawler policies both reduce what AI systems can verify. Neither is wrong on its own, but each should be a documented decision that weighs security against being described accurately in AI answers.
Cohort
Anthropic (anthropic.com): 20. Anthropic's site is open and technically clean, earning full marks on foundations. It is a well-built conventional website, not one written for machines: there is no AI context file and no structured company identity.
OpenAI (openai.com) - not scored. Our scanner was redirected and then challenged at the edge, so no reliable public evidence could be collected this month. We do not estimate scores for sites we cannot read.
Google (google.com): 19. google.com works as a search tool, not a corporate brand page. It carries almost no machine-readable description of Google as a company, and leaves that job to sub-brands and external profiles.
Meta (meta.com): 20. Meta champions open AI models, but its corporate site follows a conventional pattern: crawlable and well built, yet not described for AI systems.
SpaceX (spacex.com) - not scored. The site publishes neither a robots.txt nor an llms.txt file, so there is no stated policy for AI crawlers at all. Access is left to each crawler's default behaviour.
