Primus
(News) 70B Primus models: https://huggingface.co/collections/trend-cybertron/llama-primus-nemotron-70b-6805fd8a55d3c792e1bc1ec9
Paper • 2502.11191 • Published • 8Note Start by reading the 🚀Primus Paper! To the best of our knowledge, we are the 🏄🏽♂️ first to release datasets covering cybersecurity pretraining, IFT, and reasoning distillation. Of course, we are also the first to pretrain an LLM on a large-scale cybersecurity corpus.
trend-cybertron/Llama-Primus-Base
Text Generation • Updated • 3Note Based on Llama-3.1-8B-Instruct, continually pretrained on 2.77B tokens of cybersecurity text, achieving a 🚀15.88% improvement in the aggregated score across multiple cybersecurity benchmarks.
trend-cybertron/Llama-Primus-Merged
Text Generation • Updated • 2Note Instruct Model! While maintaining nearly the same instruction-following capability as Llama-3.1-8B-Instruct, achieving a 🚀14.84% improvement across multiple cybersecurity benchmarks.
trend-cybertron/Llama-Primus-Reasoning
Text Generation • Updated • 2Note Distilled on reasoning and reflection data from o1-preview for cybersecurity tasks, achieving a 🚀10% improvement on CISSP.
trend-cybertron/Primus-Seed
Updated • 42Note Includes high-quality cybersecurity texts manually collected from reputable sources such as wikipedia, MITRE, cybersecurity company websites, CTI, and more.
trend-cybertron/Primus-FineWeb
Updated • 14 • 2Note Includes 2.57B tokens of cybersecurity texts filtered from FineWeb.
trend-cybertron/Primus-Instruct
Updated • 41 • 2Note Includes approximately 1K QA pairs covering common cybersecurity business scenarios.
trend-cybertron/Primus-Reasoning
Updated • 43 • 1Note Includes reasoning and reflection data generated by o1-preview on cybersecurity tasks for distillation.