Arcos

Built for the most important deals: Arcos Law’s On-Premise AI with NVIDIA Nemotron

By Kiran Singh

Arcos Law is altering the way legal services are provided to capital allocators; it is a firm of a new type that combines cutting-edge technology with experienced partners in order to offer top-quality legal services to leading financial institutions.

We have always held the belief that technology should enable lawyers to act as genuine strategic advisors. Our proprietary platform carries out complicated and repetitive tasks with accuracy, which allows our team to focus on important work such as strategic analysis, thorough negotiation, and giving expert advice on key decisions.

Clients realize that traditional big law is structurally flawed—marked by high expenses, inefficiency, and a failure to keep up with the speed of modern business. Arcos Law offers an improved method which is reliable, transparent and oriented towards achieving results. Through the use of a fee structure based on value rather than the number of billable hours, we make sure that our objectives correspond to those of our clients. Our dedication to transparency, together with the use of technology that is incorporated right into the working process, means that our services remain economical and are always consistently innovative.

We operate as a hub for legal pioneers—strategic advisors dedicated to a more efficient model of practice. Our firm provides experts with proprietary tools that amplify their unique methodologies, accelerating deal velocity and sharpening analytical depth during high-stakes transactions. Fundamentally, we are technology-enabled strategists focused on serving as your most reliable legal ally.

From the very beginning, Arcos Law has firmly believed in the ability of open source and open-weight models to transform the legal industry. In the high-stakes world of large mergers and acquisitions transactions, security, privacy, and confidentiality are essential. We often work with highly sensitive Material Non-Public Information (MNPI), which demands the most rigorous confidentiality measures. By maintaining full control over our model deployments, we are able to run them on-premise or in isolated environments, thus reducing our clients’ concerns about possible leaks or unauthorized exposure of their data. Moreover, selecting open-weights gives us considerably more control over the performance, customization, and the entire engineering lifecycle of our systems. The structural flexibility this provides enables us to use smaller models that are highly finely tuned for specific corporate legal tasks—eventually leading to an improvement in the precision, accuracy, and reliability of our systems in a strict legal context when compared with general-purpose public models that depend on third-party APIs.

Integrating the Nemotron model series has greatly enhanced our technology stack. These models execute our specialized legal extraction and analysis operations with greater efficiency and significantly higher speed than the frontier model we formerly deployed.

Metric Nemotron
3.5 Lightning
Comparable
Open Weights Model
Production
Frontier Model
Benchmark Score1 65.1 63.8 60.3
Throughput (tok/s) 4478 2332 N/A2

1 Internal legal benchmark; scored out of 100, with higher scores indicating better performance.
2 Exact capacity for this external model is not publicly disclosed.

This level of efficiency directly aligns with our aim of having computers manage routine duties flawlessly, thereby allowing our human lawyers to concentrate completely on high-impact strategic advisory work.

A core feature of our automated workflows is the calculation of a number of complex risk and compliance figures for each individual provision in a contract. This thorough validation procedure is triggered automatically each time one of our lawyers makes an edit or amendment to a provision in a contract. Although this real-time system is very useful for ensuring continuous supervision, it is (given the size of our transactions) extremely computationally intensive to keep running at low latency.

Running our own custom, domain-adapted version of Nemotron 3.5 Nano has enabled a much greater degree of flexibility, cost-efficiency, and rapid performance for our entire team. By integrating NVIDIA’s advanced architecture directly into our proprietary legal workflow, we successfully deliver a modern, transparent approach to legal services that remains ahead of the curve.