LightEval
Language-model evaluation tooling for tracking prompt construction, scoring choices and evaluation configuration.
Research, software, data and methods. Inspect the context, compare your options, and take a useful next step.
Language-model evaluation tooling for tracking prompt construction, scoring choices and evaluation configuration.
Panel, instrumental-variable, and system regression estimators for econometrics.
Model-provider interface and proxy tooling for inspecting request normalization, usage accounting and backend differences.
Code-model benchmark repository with time-aware task collections for studying evaluation provenance and data contamination.
C/C++ language-model inference implementation for studying quantization and execution tradeoffs across supported hardware.
Nonlinear model fitting with named parameters, constraints, and uncertainty-analysis tools.
Analysis of molecular-simulation trajectories and coordinate data.
Molecular dynamics trajectory loading and analysis routines.
Agent-based simulation components and model-analysis tools for Python.
Compact software engineering agent implementation for examining minimal agent loops and inspectable action traces.
Experiment and model lifecycle tooling for preserving parameters, artifacts and evaluation lineage across runs.
Processing and analysis tools for electrophysiological recordings such as EEG and MEG.
Recorded license evidence is scoped to each source; it is not a blanket permission or a verification result.