2026-08-30
Temperature 0 Doesn't Buy You Reproducibility
Setting temperature to zero feels like determinism. It isn't — and in GxP work the difference will find you during an audit, not during development.
Topic
7 posts
2026-08-30
Setting temperature to zero feels like determinism. It isn't — and in GxP work the difference will find you during an audit, not during development.
2026-08-30
A tiered survey of LLM evidence in clinical trial statistical programming: benchmarked results, promising single-team studies, vendor hype, and open gaps for 2026–2027.
#llm#clinical-trials#statistical-programming#survey#benchmarks#gxp
2026-08-20
Free-form agent loops break down in GxP clinical programming. Structuring the work as a typed process DAG makes LLM agents reliable, replayable, and auditable.
2026-08-19
A GRADE-rated review of 2020–2025 automation evidence in statistical programming: real gains, mostly Low to Very Low quality evidence.
#statistical-programming#automation#evidence-quality#clinical-trials
2026-08-16
A bridge map, typed IR, and orchestrator wrap a legacy SAS TFL library unchanged — AI-ready JSON on day one, 80%+ cell-level parity, optional 92% code cut.
#sas#legacy-modernization#clinical-trials#llm-integration#statistical-programming
2026-08-13
ClinAgent splits clinical programming capability into five layers, keeping MCP tools stateless and packaging domain expertise as testable skills.
#llm-agents#mcp#clinical-trials#statistical-programming#agent-architecture
2026-08-10
Schema-only synthetic ADaM generation plateaus at 0.45 overall quality; enriching schemas from protocol/SAP/CRF knowledge graphs plus templates reaches 0.70.
#synthetic-data#adam#knowledge-graphs#llm#statistical-programming