Tweet by badlogicgames

April 26, 2026

It would be lovely if inference engines would have endpoints that advertise: - all locally cached models and their specs (input modalities, context window size, thinking, tool caling) - loaded model(s) that would allow harnesses to easily enumerate things dynamically, instead of having users (or their agents) write silly config files.

Author
badlogicgames
Date
April 26, 2026