Web Researcher (AI Benchmarking)

🏢 Gramian Consulting Group
📍 EgyptContractRemote
📅 Posted: 1w ago🔄 Updated: 1w ago
CV%
✨ AI Summary
Gramian Consultancy is seeking Web Researchers for AI Benchmarking to design challenging open-web research problems for evaluating advanced AI browsing agents. The role involves starting from verifiable facts, constructing difficult research questions, and documenting an auditable evidence trail. Responsibilities include researching across various sources, producing evidence trails, testing research question difficulty, and refining questions for clarity and difficulty. This is a short-term freelance/contractor assignment requiring a 40-hour work week and overlap with PST.
Required Skills
Information Technology
REST APIHive
Other
structured data formatsevidence tracingbenchmark constructiongovernment databasesinstitutional sourcesregistriesPDF documents
Finance, Legal & Governance
Investigations
Engineering, Construction & Trades
Validation
Soft Skills & Professional Competencies
Research
Requirements
Background in reference librarianship, archives, or special collections.Experience in investigative journalism or professional fact-checking.Experience with OSINT, due diligence, KYC, or investigative research.Background in patent, prior-art, or legal-discovery research.Experience with genealogy or historical records research.Experience with competitive quizzing, puzzle design, or puzzle-hunt construction.Familiarity with JSON or structured data formats.
Description
About UsGramian Consultancy is a boutique consultancy specializing in IT professional services and engineering talent solutions. With a strong background in software engineering and leadership, we help companies build high-performing teams by matching them with professionals who truly fit their needs.Role overviewOur client is building advanced evaluation benchmarks for frontier AI browsing agents.We are looking for highly skilled investigative researchers who can design research problems that remain difficult even for state-of-the-art AI systems with full web access. This is not a traditional subject-matter or content-writing role. The work is centered on open-web investigation, evidence tracing, source validation, and benchmark construction.You will begin from a verifiable fact, work backwards to construct a difficult research question, and document a complete, auditable evidence trail showing how the answer can be independently verified.CONTRACT: Short-term freelance/contractor assignment, 8 weekaCOMMITMENT: 40 hours per week, at least 4 hours per day, including 4 hours overlap with PSTLOCATIONS: Remote, GLOBALResponsibilitiesDesign challenging open-web research problems for evaluating advanced AI browsing agents.Start from objectively verifiable facts and construct questions that make those facts difficult to discover.Build multi-step clue structures involving dates, people, places, organizations, works, events, records, and quantities.Ensure every clue is independently verifiable and subject to clear constraints.Research across government databases, institutional sources, archives, registries, and PDF documents.Produce complete evidence trails citing exact pages, tables, sections, or records.Record validation searches and document what obvious search approaches return.Test whether research questions remain difficult across multiple search attempts.Refine questions to eliminate ambiguity while preserving difficulty.Deliver structured, reproducible research documentation.
✨ Premium Match Details
Deep-dive CV analysis, customized Cover Letters, and Interview prep!
📊 Match Analysis
Insights against your active CV
📊
Personalized Match Analysis
Upload your CV to see exact matching percentages, detailed skills mapping, and gap analysis for this role.
🎯 Overalli74%
⚡ Skillsi85%
View Breakdown
Ontology Match: 85.0
Matched:✓ Requirements Matching✓ Ontology Skills Mapping
📜 Eligibilityi49%
View Breakdown
Local: 19600%
🏗️ Career Fiti91%
View Breakdown
Seniority: 91.0
📋 Requirementsi67%
View Breakdown
Domain: 67.0
🔥 Motivationi78%
View Breakdown
Title Fit: 78.00