We benchmarked 4 @DSPyOSS adapters, 2 modules, and 5 models on a real structured extraction task. The results were not what we expected.
𝗧𝗵𝗲 𝘁𝗮𝘀𝗸: Extract a list of typed @pydantic objects (decisions and learnings with literal fields, optional values, nested lists) from