{"categories":[{"label":"Artificial Intelligence","url":"https://skillfed.io/packages/category/scientific-engineering-artificial-intelligence"}],"enrichment":{"capability":"Inspect is a framework for evaluating large language models, providing built-in components for prompt engineering, tool use, multi-turn dialogue, and model-graded scoring across any model.","skillfed_tags":["llm-evaluation","model-testing","ai-security"],"use_cases":["Run standardized evaluations against multiple LLM providers to compare model performance on safety, reasoning, and capability benchmarks","Build custom evaluation pipelines combining prompt engineering, tool use, and model-graded scoring for domain-specific tasks","Integrate LLM evaluation into CI/CD workflows to validate model behavior before deployment","Develop and share new evaluation techniques as Python extensions compatible with the Inspect framework"],"what_it_does":"Inspect is a framework for evaluating large language models, created by the UK AI Security Institute. It provides a structured environment for running evaluations on any model, with built-in support for prompt engineering, tool use, multi-turn dialogue, and model-graded scoring. The framework includes over 200 pre-built evaluations ready to run out of the box.\n\nThe package has a substantial dependency graph (40 runtime dependencies including pydantic, fastapi, boto3, numpy, and AWS integration libraries) reflecting its role as a full-featured evaluation platform. It supports both Python development workflows and web-based UI interaction through a TypeScript/React frontend. Active maintenance and recent releases indicate ongoing development.","worth_installing":"Yes, if you need to evaluate LLMs systematically. The framework is actively maintained, permissively licensed, and backed by a government security institute. The 40 runtime dependencies and Python 3.10+ requirement are substantial but typical for a full-featured evaluation platform. Install if you're building evaluation pipelines or comparing model outputs; skip if you only need lightweight model testing."},"id":"inspect-ai","links":{"html":"https://skillfed.io/packages/inspect-ai","md":"https://skillfed.io/packages/inspect-ai.md","pypi":"https://pypi.org/project/inspect-ai/"},"maintenance":{"status":"active"},"meta":{"latest_release":"2026-08-12","license_spdx":null,"license_treatment":"permissive","name":"inspect-ai","python_support":"supports_current","summary":"Framework for large language model evaluations"},"popularity":{"monthly_downloads":12105306,"position":1342,"tier":"top_5000"},"security":{"n_vulnerabilities":0},"version":"0.3.258"}
