# Fig.inc > Fig is a frontier AI lab pioneering systems capable of continual learning and autonomous action across complex, open-ended environments. Public Ghost content for AI and LLM tooling. Use `/llms-full.txt` for consolidated page and post context. Append `.md` to any post or page URL to get the content in Markdown (for example, `/example-post.md`). ## Pages - [Privacy Policy](https://www.fig.inc/privacy.md) - Last updated: June 15, 2026 1. Introduction Metarch Corporation ("Fig," "we," "us," or "our") is an artificial intelligence research company. This Privacy Policy (the "Policy") describes how we collect, use, and disclose your information when you access or use our website and blog at fig.inc, or ot… ## Posts - [Fixing Failures in Browser-Use Models: Why More Data Isn't Enough](https://www.fig.inc/blog/fixing-failures-in-browser-use.md) - [GUI-Perturbed: Breaking Browser-use Models using Domain Randomization](https://www.fig.inc/blog/gui-pertubed-breaking-browser-use-models.md) - Original Model Result ✓ Correct ✕ Precision Variant (70% zoom) Model Result ✗ Clicked on fake 'View Deal' ad ✕ Figure 1: Sample: 21 of 390: "Click on 'View Deal' button for flight '#2125, #2126'", Direct Instruction, No Reasoning. Result: UI-TARS1.5-7B clicked on the fake 'view deal' button in the… - [Domain Randomization for Computer Control](https://www.fig.inc/blog/domain-randomization-for-computer-control.md) - Yangyue Wang1, 2, Harshvardhan Sikka1, 2, Yash Mathur*2, Tony Zhou*2, Jinu Nyachhyon*2, Pranav Guruprasad1, 2 * Equal contributions. 1Fig; 2Manifold Research Group. 0:00 /0:09 1× TL;DR * GUI models scoring 90%+ on standard benchmarks fail under basic visual variations like a 70% browser zoom. Curre… - [The Scaling Laws Are Breaking](https://www.fig.inc/blog/the-scaling-laws-are-breaking.md) - Part 1 of a series examining the need and search for new scaling laws. This piece reflects ongoing research at Fig. We welcome discussion and collaboration as we work to formalize these observations. OpenAI's Orion consumed 15-30× more training compute than GPT-4. By the scaling laws that have guid… - [To Solve the Benchmark Crisis, Evals Must Think](https://www.fig.inc/blog/to-solve-the-benchmark-crisis-evals-must-think.md) - GPT-4 scored 95% on HumanEval. So did Claude. So did Gemini. But your production deployment still breaks on basic customer queries. We've collectively entered the what is fast becoming a dangerous phase of AI development: when benchmarks tell us nothing about what actually matters. Models have memo… ## Optional - [RSS Feed](https://www.fig.inc/blog/rss/) - [Sitemap](https://www.fig.inc/sitemap.xml) - [Full content of pages and posts](https://www.fig.inc/llms-full.txt)