Skip to content
    HAQQ
    • Pricing
    Get Started Free
    Get Started FreeBook a Demo
    Log in

    Justinian

    The proprietary Legal AI engine behind HAQQ Chat, eFirm and the mobile app - built for lawyers and for everyone who is not one.

    Test JustinianSee the benchmark
    ArchitectureCapabilitiesEvaluationLimits & safetyResearch
    Justinian establishes what you are actually asking, which jurisdiction governs it, and what the law says, before a single word is drafted. Built for work that has to be reviewed, cited and signed.
    Jad JabbourAI Engineer at HAQQ Legal AI
    Every legal claim
    CitedEvery legal claimGrounded in retrieved law. Where the law is unsettled, Justinian flags it rather than resolving it for you.
    Major practice areas
    24Major practice areasDrafting, review, research, cross-border analysis and study - one engine, not five tools.
    Highest score of 19 models
    49 / 50Highest score of 19 modelsTop score on our published cross-model benchmark, one part of the 250 complex legal tasks Justinian is evaluated on.
    Long-horizon tasks
    1,372Long-horizon tasksWhole matters run end to end and graded all-pass: every step has to pass, partial credit does not count.

    Everything behind a single answer

    One answer runs on a full stack: product surfaces, security and governance, frontier models, grounded data, and jurisdiction-aware reasoning across MENA, Europe and the US.

    The Corpus Juris Civilis: its three parts, the Codex of 529 and the Digesta and Institutiones of 533, shown as an opened codex with rubricated initial lines.

    Why Justinian?

    The name Justinian honors Emperor Justinian I of the Eastern Roman Empire - the architect of history's most influential legal codification. Where Emperor Justinian unified Roman law for an empire, our Justinian engine unifies legal intelligence for the practitioner.

    “The empire's legal system needed repair. There existed three codices of imperial laws and other individual laws, many of which conflicted or were out of date.” - On Emperor Justinian's mandate, 527 AD

    The rest of the legal-AI landscape

    We map every legal AI vendor we know of, by category, including the ones that beat us on individual tasks.

    point tool to full-stack platformgeneral-purpose to legal-specializedthe full-stack legal OS
    [GENERIC AI]General-Purpose LLMs
    [LEGALLY-TRAINED LLMS]Legally-Trained LLMs
    [LEGAL-AI TOOLS]Legal AI Point Tools
    [PRACTICE MANAGEMENT]Traditional Practice Management
    [LEGAL AI PLATFORM]HAQQ
    [CONTRACT DRAFTING]Contract Drafting & CLM
    [AI-NATIVE FIRMS]AI-Native Law Firms
    [SKILL REGISTRIES]Community Legal Skill Registries
    [THE INCUMBENTS]The Research Incumbents
    [CLOSET ADOPTERS]The Closet Adopters
    [THE EVALUATORS]The Evaluators

    Eleven categories, one map. HAQQ is the only one in the full-stack, legal-specialized corner. Click any node to drill in.

    See the landscape

    Limitations

    Despite the leap in capabilities, the engine exhibits several limitations common to legal generation engines. Understanding these boundaries is essential for responsible use.

    • Input Dependency

      Incomplete or incorrect facts produce incomplete or incorrect analysis. The engine cannot independently verify factual claims - quality in determines quality out.

    • In
      Not connected

      Knowledge Boundaries

      Cannot access information outside training data and connected sources. Real-time legal updates require explicit integration with your firm's data.

    • Professional Review Required

      Every output requires review by a qualified professional. This is not a limitation to be solved but a design principle.

    Safety & Professional Responsibility

    • Disclosure

      AI-generated content is always labelled, so you always know what came from the engine.

    • Competence

      It works inside legal domains and refuses to speculate outside them.

    • Confidentiality

      Zero data retention, end-to-end encryption, and no training on your data.

    • Oversight

      Every output needs professional review before it reaches a client or a court.

    Legal responsibility always remains with the lawyer. This is intentional.

    Read our security practices

    AI Research Behind Justinian

    Deep dives into the architecture, experiments, and benchmarks that power our legal AI engine.

    1. Measured

      Top legal tasks by AI usage

      The legal tasks lawyers most often hand to AI, ranked by share of usage.

      View the data
    2. Measured

      Cross-model legal benchmark

      Every legal AI model we tested, scored out of 50 across legal task categories.

      View the data
    3. Measured

      AI autonomy scores by task

      How autonomously AI can handle each legal task, and where it must stop.

      View the data
    4. Measured

      Tasks that still need a human

      The share of each legal task professionals say still requires human judgment.

      View the data
    5. Measured

      Time savings by legal task

      Median hours saved per task when a lawyer works with AI rather than alone.

      View the data
    6. Illustrative

      Risk and hallucination rates

      Estimated error rates across legal tasks, highest for citation and jurisdiction work.

      View the data
    Browse the full research hubRead the HAQQ Legal AI Index

    Frequently Asked Questions

    How is Justinian different from ChatGPT?+

    ChatGPT is a general-purpose language model that generates plausible-sounding text. Justinian is a legal AI engine: before any general model is called, it establishes what you are actually asking, which jurisdiction applies, what language the governing law is in, and what the request needs. The general model then runs on a question that has already been framed. It applies legal rules to facts, respects jurisdictional boundaries, cites retrieved sources, and produces structured, client-ready deliverables.

    Does Justinian cite real case law?+

    Yes. Justinian searches verified legal databases and authoritative sources before answering. Every citation is traceable and verifiable. When sources conflict or the law is ambiguous, Justinian flags the uncertainty rather than presenting a confident but wrong answer.

    Can Justinian draft contracts in Arabic?+

    Yes. Justinian drafts natively in Arabic, English, French, and other languages with full RTL support. It understands legal nuance in each language and can produce bilingual contracts while maintaining legal precision across both versions.

    What jurisdictions does Justinian support?+

    Justinian reasons across MENA, Europe, the US and international arbitration frameworks including ICC and UNCITRAL, and it drafts in English, Arabic and French. Depth varies by jurisdiction, because it depends on how well the source law is digitised and how much of it is public. Rather than quote a single global number, we would rather answer the question for the specific jurisdiction you work in.

    Is Justinian's reasoning auditable?+

    Yes. Every Justinian output includes a transparent reasoning chain - you can see which rules were applied, which sources were consulted, and how the conclusion was reached. This auditability is critical for professional accountability and client trust.

    What happens before Justinian answers a question?+

    Your question does not go straight to a general-purpose model. It first passes through a step we own and control, which establishes what you are actually asking, what kind of legal work it is, which jurisdiction is in play, what language the governing law is written in, and the facts inside your request as structured data. Only then does the general model run, and it arrives with the question already framed rather than guessing at it.

    Why does doing work before the model call make the answer better?+

    Two reasons. A model's context window is finite, so anything loaded unnecessarily competes for attention with your actual document. And a model handed structured facts instead of a wall of prose spends its budget on legal reasoning rather than on parsing what it was given. The second effect is larger than most people expect, and it is why the same general-purpose model produces better legal work inside Justinian than inside a simple wrapper.

    Which AI model does Justinian use, and are you locked into one vendor?+

    We deliberately do not build on a single vendor. A third-party frontier model performs the final generation, but every layer that carries our value runs independently of any particular one: orchestration, tooling, retrieval, memory and safety. Justinian is architected to route different kinds of legal work to different models, so if a better or cheaper model appears, adopting it is a configuration change rather than a rebuild.

    Can Justinian read scanned documents and photographs?+

    Yes. We treat document intake as a first-class engineering problem rather than a preprocessing afterthought, because a large share of legal work arrives as a scan, a photograph of a stamped page, or a PDF that was printed and re-scanned crooked. An error introduced at that stage cannot be recovered later: the system would reason impeccably about the wrong clause number.

    Does Justinian remember my previous matters?+

    Yes. Justinian holds the relationships between your matters over time and draws on them when a question calls for it, rather than simply replaying recent messages back into the prompt. That is what lets it recognise that the counterparty in a contract you are reviewing today is the same one from a dispute two months ago.

    Does Justinian hallucinate?+

    Any system built on a language model can produce a wrong answer, and any vendor who tells you otherwise is not being careful with words. What good architecture does is make wrongness rarer and detectable: legal claims are generated against retrieved sources, every claim carries a citation you can open, and where sources conflict or the law is genuinely ambiguous Justinian flags the uncertainty instead of resolving it for you. It is built to be reviewed by a professional, not to replace one.

    How do I tell whether a legal AI product is just a ChatGPT wrapper?+

    Four questions in a demo, and they work on any vendor including us. Ask it something outside law: a wrapped general model will usually answer a medical question. Ask the same legal question in two languages: if the substance changes, language is handled after the reasoning rather than before it. Upload a badly photographed document. And push it toward an unsettled area of law, where a system built for professional work will tell you the ground is uncertain instead of producing a confident answer.

    How is Justinian evaluated?+

    Against legal-specific benchmarks rather than general ones, because a model that scores well on general reasoning can still be useless on a clause. Clause-level validation, jurisdictional accuracy and continuous feedback from practising lawyers drive what the engine does before and around every model call: what gets retrieved, which capabilities run, and what gets flagged for human review. The results are published rather than described, including the ones where a frontier model beats us.

    How do I know a number on this page is real?+

    Every figure we publish declares how we know it. Measured means it comes from our own usage data or our own benchmark run. Cited means it is a third-party finding, linked to its source. Illustrative means it is a directional scenario, and it carries a badge saying so, including on the worked examples further up this page. If a number does not say which of the three it is, treat that as a defect and tell us.

    Do you publish results where another model wins?+

    Yes, and you should be suspicious of any legal AI vendor who does not. A benchmark that its author always wins is marketing, not evaluation. Our published runs include tasks and categories where a frontier model scores above us, because a buyer needs to know where the tool is strong and where it is not before they rely on it in front of a client.

    Try Justinian

    Experience a legal AI engine built at the frontier - and constrained by law.

    HAQQ across all devices
    HAQQ Legal AI Platform Logo

    Your Legal AI Twin & Practice Management System for drafting, billing, and winning.

    Download on theApp StoreGet it onGoogle Play

    Documentations

    • Docs opens in a new tab
    • Getting Started opens in a new tab
    • Newsroom opens in a new tab
    • Product Updates opens in a new tab
    • Status opens in a new tab
    • Security
    • FAQ opens in a new tab
    • Community opens in a new tab
    • Support opens in a new tab

    Academy

    • Partner opens in a new tab
    • Course opens in a new tab
    • Skills opens in a new tab
    • Clause opens in a new tab
    • Prompt Library opens in a new tab
    • Tools opens in a new tab
    • Research Hub opens in a new tab
    • Documents opens in a new tab

    Website

    • eFirm
    • Legal AI Chat
    • Mobile App
    • Justinian AI Engine
    • HAQQ eBar
    • HAQQ eWallet
    • Pricing
    • Compare Us
    • Solutions
    • Blog
    • Meet Team
    • Join Us opens in a new tab
    Open App
    • Localesar en fr es it de pt
    • Contactinfo@haqq.ai
    • Statusoperational·grounded
    • Terms of Service
    • Privacy Policy
    • Cookie Policy
    • Data Processing
    • humans.txt opens in a new tablawyers.txt opens in a new tabsecurity.txt opens in a new tab
    © 2026 HAQQ Inc. All rights reserved.Product engineered in-house by HAQQ. Website built with modern web tools.
    User Query

    Reviewthesecontractsforcompliancegapsacrossourjurisdictions.

    PDFDOCDOCXXLSDataroom+17 formats
    🇦🇪UAE🇸🇦Saudi Arabia🇪🇬Egypt🇫🇷France
    Decided here04Entities

    Product & Interfaces

    One platform, every surface

    Try it
    HAQQ ChatHAQQ eFirmMobile AppeBarClient PortalAPI / MCP
    Decided here10Thread title
    Security & Governance
    Cross-Matter IsolationAudit TrailsRBACAnonymization
    Decided here09Scope and safety

    JUSTINIAN AGENTIC HARNESS

    The Ontology of Intelligence

    Resolved before the model call

    01Prompt rewrite02Intent07Model routing08Reasoning depth

    10/10 decisions resolved

    Architected to route acrossOpenAI model logoAnthropic model logoGoogle model logo

    Then the agentic loop

    Sub-Agent OrchestrationMemoryDeep ReasoningLookahead + RollbackCitation AuditHuman-in-the-Loop

    Context & Knowledge

    Your firm's institutional memory

    RAG
    PlaybooksPoliciesMemosTemplatesPrecedentsClause LibrariesMatter History
    Decided here05Thread state

    Data & Integrations

    Live legal sources

    Secure
    Web SearchStatutesCase LawGazettesRegulatory DBsUSPTOEUIPOWIPO

    Firm

    Your legal operating system

    MattersContractsDocumentsHearingsTasksCalendarFinancialKYCLeadsEmails

    Jurisdictions

    Multi-region coverage

    MENAEUUSASouth AmericaAsiaAfrica
    🇦🇪 UAE🇸🇦 Saudi Arabia🇶🇦 Qatar🇴🇲 Oman🇯🇴 Jordan🇪🇬 Egypt🇱🇧 Lebanon🇮🇶 Iraq🇸🇾 Syria🇲🇦 Morocco🇹🇳 Tunisia🇩🇿 Algeria🇫🇷 France🇪🇸 Spain🇮🇹 Italy🇵🇹 Portugal🇧🇪 Belgium🇳🇱 Netherlands🇸🇪 Sweden🇬🇷 Greece🇹🇷 Turkey🇺🇸 US🇨🇦 Canada🇧🇷 Brazil🇲🇽 Mexico🇮🇳 India🇸🇬 Singapore🇳🇬 Nigeria🇰🇪 Kenya🇿🇦 South Africa
    Decided here03Language

    Legal Capabilities

    DraftingRedliningDue DiligenceCompliance ReportsLitigationAnalysisCase SummariesOffer LettersDeposit DisputesWillsTenant RightsTraffic FinesFreelance ContractsPrenupsVisa GuidanceGDPR ChecksCross-Border DealsTrademark ChecksEmployment Disputes
    Decided here06Capability selection
    Client-Ready Output

    Why Justinian is different

    A proprietary, in-house legal AI engine, not a wrapper around a public model. Ten decisions are resolved before a frontier model is ever called.

    Your jurisdiction
    Outside

    Sovereign AI & On-Premise Deployment

    Deploy on-premise, private cloud, or air-gapped environments. Full data sovereignty with no data ever leaving your jurisdiction.

    Built-In Anonymization & Encryption

    Proprietary anonymization layer strips PII before inference. End-to-end encryption at rest and in transit. GDPR, SOC 2, ISO 27001 compliant.

    One question
    🇦🇪
    🇫🇷
    🇺🇸

    Multi-Jurisdiction Legal Reasoning

    Jurisdiction-aware analysis grounded in local statutes, case law, and regulatory frameworks across 200+ legal systems.

    1
    2
    3
    ?
    Sources

    Knowledge-Graph Grounded Outputs

    Proprietary legal ontology, clause-level analyzers, and constraint engines ensure every output is legally defensible.

    ChateFirmMobileeBarAPI
    Justinian

    Full-Stack Legal OS

    Not a wrapper. A single proprietary engine powering AI chat, practice management, eBar, payments, and mobile with no third-party AI dependencies.

    Justinian orchestration
    any frontier model
    Same output contract

    LLM-Agnostic & Future-Proof

    Swap underlying models without re-engineering. Proprietary orchestration layer ensures consistent quality across any frontier LLM.

    Benchmarked across major practice areas

    Justinian against the frontier models on legal, tax, journalism, general and safety work. Scores are out of 100, run 2026-08. The best score in each row is bold. An empty cell means the model was not run on that benchmark: it is not a zero, and it is left out of that model's averages.

    Benchmark
    Overall average83.279.578.578.077.276.575.775.773.673.071.9
    Legal
    Stanford LegalBench85.781.882.384.382.982.383.381.482.878.876.8
    Info. Retrieval57.251.253.653.853.653.853.047.951.249.051.5
    Reasoning78.275.273.277.873.168.474.873.170.866.570.1
    Classification74.770.570.974.270.270.071.670.469.868.767.8
    Doc. Processing & RAG83.278.979.778.276.578.679.074.771.274.173.7
    Summarisation90.088.390.089.490.086.387.784.887.982.589.1
    Contract Under.77.474.471.075.973.674.977.269.671.768.468.3
    Human Queries89.785.489.286.487.388.661.484.387.186.786.4
    Deep Research90.990.888.980.689.085.989.886.078.487.384.0
    Harvey LAB86.886.985.755.584.676.183.780.956.370.683.1
    Legal average81.478.378.475.678.176.576.175.372.773.375.1
    Tax
    Deep Research86.385.882.476.083.580.985.182.070.478.780.7
    Tax Q&A88.986.987.984.585.786.688.388.788.285.182.8
    Tax average87.686.385.180.284.683.886.785.479.381.981.8
    Journalism
    Deep Research84.682.380.983.079.066.384.574.176.277.278.6
    Journalism average84.682.380.983.079.066.384.574.176.277.278.6
    General
    Factuality82.971.473.782.666.268.369.559.871.776.269.1
    Long Context75.675.275.375.075.970.073.570.774.653.069.5
    Multilingualism85.883.278.485.781.782.784.979.574.277.678.4
    Instr. Following91.786.191.484.889.489.085.886.887.385.689.2
    Writing80.779.380.378.578.179.178.278.779.977.980.7
    Reasoning76.273.768.474.867.371.676.266.866.866.665.0
    General Agent89.283.489.184.477.575.661.074.187.185.572.0
    Coding66.857.439.950.056.050.866.857.440.943.945.6
    Maths98.798.794.097.491.297.796.285.995.594.994.9
    General average83.178.776.779.275.976.176.973.375.373.573.8
    Safety / Values
    Political Neutrality97.982.897.393.391.083.833.582.385.351.531.3
    Robustness78.578.260.365.350.468.372.777.142.266.738.1
    Adversarial Testing95.9—93.4—93.6—87.3—78.895.981.1
    Safety / Values average90.8—83.6—78.3—64.5—68.771.450.2
    Full roster: 11 scored, 13 queued, 5 without evaluation access
    • HAQQ (Justinian) (HAQQ) scored
    • Claude Fable 5 (Anthropic) queued
    • Claude Opus 4.8 (Anthropic) scored
    • DeepSeek V4 Pro (DeepSeek) scored
    • GPT-5.6 Sol Pro (OpenAI) queued
    • Gemini 3.1 Pro (Google) scored
    • Legora (Legora) no evaluation access
    • Grok 4.5 (xAI) queued
    • Spellbook (Spellbook) no evaluation access
    • Clio Duo (Clio) no evaluation access
    • Llama 4 Maverick (Meta) queued
    • Mistral 3 (Mistral) queued
    • Qwen3.7 Plus (Alibaba) queued
    • Kimi K3 (Moonshot AI) scored
    • Mike OS (Mike) queued
    • GPT-5.6 Luna (OpenAI) queued
    • Claude Haiku 4.5 (Anthropic) queued
    • GLM 5.2 (Z.ai) scored
    • Claude Sonnet 5 (Anthropic) scored
    • Nova 2 Lite (Amazon) queued
    • Thomson 1.0-Large (Thomson Reuters) scored
    • Harvey Tenet (Harvey) no evaluation access
    • LexisNexis +AI (LexisNexis) no evaluation access
    • Gemini 3.5 Flash (Google) queued
    • Nemotron 3 Ultra (NVIDIA) queued
    • DeepSeek-R1 (DeepSeek) queued
    • Qwen3.5 397B (Alibaba) scored
    • GPT 5.4 (OpenAI) scored
    • Snowdon 1.0-Large (Snowdon) scored

    Queued models are reachable over an API and are a scheduling question. The rest publish no evaluations and offer no self-serve access, so scoring them needs their cooperation. No score is estimated: a cell is either measured or it is absent.

    Overall average is the mean of the 24 benchmarks excluding adversarial testing, which is the study's own method: it reproduces every published overall exactly, and four models were never run on that benchmark. Domain averages do include it.

    Frameworks this run draws on: Stanford LegalBench, Harvey LAB, VLAIR (Vals AI), ALARB, BigLaw Bench, CUAD, LegalCiteBench, LawBench / LexEval, Legal Benchmarks.