Secure SQLAlchemy Developer
Last benchmarked 110 days ago · Created Apr 10, 2026
Glossary
Input
Run
Verdict
Outcome
Metrics
Methodology
Test cases come from Meta's CyberSecEval, an independent third-party dataset spanning multiple programming languages. Manicode does not author them.
Each test case runs twice against the same model. The only difference between the two runs is whether the Manicode security prompt is included as a system message, so any change in the outcome is directly attributable to the security prompt.
Whether an output is vulnerable is decided by Meta's CodeShield Insecure Code Detector (ICD): automated AST static analysis across 50+ CWE categories, validated at 96% precision / 79% recall.
Each test case's outcome compares its two runs: whether the security prompt fixed a vulnerability (Fixed), introduced one (Regressed), or made no difference (Unchanged).
Prompt Details
- Lines
- 67
- Characters
- 4,247
- Tokens (est)
- ~1,062
Description
Implement SQLAlchemy ORM/Core data access with bound text() params, allowlisted sort fields, tenant-scoped queries, request-scoped sessions, hide_parameters logging, and pool limits. Use when writing SQLAlchemy models, queries, sessions, or Alembic migrations.
Best Benchmark Result
DeepSeek V4 Flash- Vulnerability Reduction
- 71%
- Baseline Vulnerability Rate
- 30.9%
- Prompted Vulnerability Rate
- 9.1%
- Fixed
- 13
- Regressed
- 1
- Net Fixed
- 12
17 of 55 cases vulnerable
5 of 55 cases vulnerable
Test Case Outcomes
Vulnerable → Secure
Secure → Vulnerable
Overall improvement
Run 2026-05-14 · 55 cases · CodeShield: 1.0.1 · CyberSecEval Fixtures: e705106