OpenMythos
- 4 Devlogs
- 50 Total hours
A cyber security research model inspired from anthropic's project glasswing.
A cyber security research model inspired from anthropic's project glasswing.
Hit 9,000+ users on OpenMythos last week
OpenMythos is a cybersecurity-focused LLM trained on ArXiv cs.CR papers and real CVE data.
If you’ve tried it:
→ Share your feedback
→ Found a vuln in your project using it? We’d love to hear
Want to stress-test your own project’s security?
Give it a shot 👇
https://huggingface.co/spaces/build-small-hackathon/OpenMythos
Hey everyone! OpenMythos benchmarks are finally here.
Sorry it took about a week to post these.
The delay was mainly because SWE-bench results weren’t matching up with Qwen 3.6 27B official numbers.
Turns out Qwen used a different eval harness and also refined/filtered the benchmark problems, even there prev 3.5 (72.4 in SWE Verified ) version benchmark score is not matching with the numbers published in 3.6 (75 in SWE Verified).
Anyway, here are the results across SWE-bench Pro, CyberGym, and cybench.
OpenMythos holds up pretty well for a small cybersecurity-focused model!
But it has capability to do better. So, will train it further.
Demo: https://huggingface.co/spaces/build-small-hackathon/OpenMythos
Model: https://huggingface.co/build-small-hackathon/OpenMythos
Generating dataset for OpenMythos