The earlier version is built and proven end to end on one customer's data. The newer version takes the same design across PEP Health's full customer base, with reasoning handled by Llama 3.3, self-hosted on PEP Health's own AWS infrastructure.
That generalised version is built and one live test away. It has been checked against a scripted test response, but has not yet run end to end against a live model server.
The 60 per cent figure remains a design target for reducing manual report-writing effort, not a measured result. It stays framed that way here.