response streaming.md

Host ai workflow runner with FastAPI: streamed responses

Summary

Deploy an AI workflow runner built with FastAPI on Ample using the streaming response service pattern. Compute runs the app in an isolated microVM behind a public HTTPS URL, with a managed PostgreSQL 16 database auto-provisioned and injected as DATABASE_URL. Verified on FastAPI: a Server-Sent Events endpoint reads incrementally through the public HTTPS gateway.

Infrastructure requirements

Prerequisites

Exact tested configuration

Steps

  1. Build and start. pip install into .ample/python from requirements.txt, then uvicorn from run.py reading PORT on the python-3.12 template. The server must bind 0.0.0.0 on PORT.

  2. Implement the pattern on PostgreSQL. The fixture's module implements streaming response service. Copy the approach into your schema; keep migrations idempotent and run them with --release-command.

  3. Deploy. Run the synchronous deploy once and read the result (exit 0 live, 1 failed, 2 blocked).

    ample deploy . --name <app-name> --public --start "python3 run.py"
    
  4. Verify. Fetch the live URL and the pattern self-test route(s) from the example; then run your own checks. On failure read ample logs <deployment_id> --kind build then --kind runtime.

    ample logs <deployment_id> --kind build
    

Tested examples

Success checks

Limitations

Cost estimate

Estimated 10.00 USD per month:

Verification evidence

Last verified: 2026-09-21T01:14:18Z

Execution binding

MCP tool ample_deploy, schema hash observed at revision.