Original task design
Each task starts with a target security skill, a clear objective, and constraints chosen for the model and agent setup.
CYVARC builds original cybersecurity task datasets for companies developing local or specialized models, frontier labs, and data-labeling providers. Tell us which capabilities matter, how your agents operate, and what your pipeline expects. We design the tasks and validation process around those requirements.
A useful security task needs a stable environment, a solution that another expert can reproduce, and a reward signal that measures the intended behavior. We treat those parts as one deliverable.
Each task starts with a target security skill, a clear objective, and constraints chosen for the model and agent setup.
A specialist solves the task from scratch so unclear instructions, brittle setup, and accidental shortcuts are found before delivery.
We inspect useful successes and failures, remove noise caused by broken environments, and prepare clean reference trajectories when requested.
Select the capabilities your model needs to learn, then tune difficulty, context, tool access, and task volume to its current level.
Commission private, high-difficulty tasks with reproducible environments, independent solution checks, reviewed runs, and reward logic tested for shortcuts.
Add cybersecurity task writing, domain review, solution validation, and reward design to an existing labeling or post-training workflow.
Request standalone task packages, reviewed traces, SFT-ready trajectories, reward checks, or a complete dataset in the schema your pipeline uses.
Share the target capability, domain, expected volume, agent setup, delivery format, and review requirements. We will shape the task and validation plan around them.