Scientific framework for iterative LLM prompt improvement using multi-dimensional scoring, threshold optimization, cross-validation, and an OPRO-style agent loop. Built on AWS Bedrock with a React + FastAPI observation GUI.
react python typescript cross-validation active-learning fastapi llm prompt-engineering aws-bedrock prompt-optimization llm-as-a-judge scientific-evaluation
-
Updated
Mar 11, 2026 - Python