Skip to main content
In-process model runtime. Loads exported .pt2 checkpoints and runs inference in the current process. This package was extracted from the FastAPI server (api/) so that both the server and the SDK in-process backend share a single implementation of model loading, preprocessing, GPU forward passes, and dynamic request batching. Public surface:
  • BaseServerModel, PreparedSample.
  • BatchScheduler.
  • build_server_model.
  • .pt2 utilities from bioptimus.runtime.pt2.