Openlayer’s Blog

No commitments, unsubscribe any time.

September 17, 2026

LLM Coding Benchmarks: The Complete Guide September 2026 (Updated)

AI agents

September 17, 2026

AI Incident Response for Model Failures (September 2026)

Governance

September 17, 2026

Continuous AI Risk Monitoring: Beyond Snapshots (September 2026)

AI evals

September 17, 2026

How to Red-Team AI Models Effectively (September 2026)

Governance

September 8, 2026

Tiering AI Systems: Internal Risk Classification (August 2026)

AI evals

September 8, 2026

AI Risk Assessment Report for Auditors (August 2026)

AI evals

July 28, 2026

NAIC AI Model Bulletin: What Insurers Must Prepare for in July 2026

AI evals

July 21, 2026

Quantify and Prioritize AI System Risk with Scoring (July 2026)

Compliance

July 13, 2026

PII Detection in LLM Outputs: AI Team Guide (July 2026)

AI evals

July 13, 2026

RAG Evaluation in Production: Groundedness, Faithfulness, and Retrieval Quality (July 2026)

AI evals

June 2, 2026

OpenAI evals: A complete guide to evaluation frameworks in March 2026

AI evals

March 30, 2026

LLM-as-judge: A complete guide to evaluation best practices in March 2026

Model quality

March 30, 2026

Model monitoring in 2026: A complete guide for ML teams

AI evals

March 27, 2026

LLM evaluation metrics: Complete guide for March 2026

AI evals

March 9, 2026

Agent evaluation: Complete guide to testing AI agents in March 2026

AI evals

February 11, 2026

RAG Groundedness Evaluation Guide (Feb 2026)

AI evals

January 29, 2026

Needle in a Haystack: AI Testing Guide (Jan 2026)

AI evals

January 2, 2026

Galileo reviews, pricing, and alternatives (January 2026)

AI evals

December 22, 2025

Best AI drift detection tools for production models (December 2025)

Observability

December 22, 2025

Braintrust reviews, pricing, and alternatives (December 2025)

Work on the future.