[NeurIPS 2026] KaliBench: A Fine-Grained Benchmark for Cybersecurity Tool Use on Kali Linux with Runtime-Free Verifiable Rewards
AI & ML interests
None defined yet.
Recent Activity
View all activity
[CVPR 2026] SPARROW: Learning Spatial Precision and Temporal Referential Consistency in Pixel-Grounded Video MLLMs
[ICLR 2026] RedSage: A Cybersecurity Generalist LLM. List of Cybersecurity Benchmarks Datasets.
Benchmarks for Evaluating Trade-offs in Image Generation (ICCV 2025, ACMMM 2026)
Visually Grounded Commonsense Reasoning Supervision for CLIP (ECCV 2026)
-
RISys-Lab/ReasonCLIP-B32-S1
Zero-Shot Image Classification • 0.2B • Updated • 4 -
RISys-Lab/ReasonCLIP-L14-224-S1
Zero-Shot Image Classification • 0.4B • Updated • 2 -
RISys-Lab/ReasonCLIP-L14-336-S1
Zero-Shot Image Classification • 0.4B • Updated • 5 -
RISys-Lab/ReasonSigLIP-So14-384-S2
Zero-Shot Image Classification • 0.9B • Updated • 2
[EMNLP 2026] SAFIRE: Safety-Critical Benchmark for Fine-grained Fire and Smoke Understanding in Multimodal LLMs
[ICLR 2026] RedSage: A Cybersecurity Generalist LLM. Continued Pretraining and Post-trained RedSage Models.
[ICLR 2026] RedSage: A Cybersecurity Generalist LLM. List of RedSage Datasets
Visually Grounded Commonsense Reasoning Supervision for CLIP (ECCV 2026)
[NeurIPS 2026] KaliBench: A Fine-Grained Benchmark for Cybersecurity Tool Use on Kali Linux with Runtime-Free Verifiable Rewards
[EMNLP 2026] SAFIRE: Safety-Critical Benchmark for Fine-grained Fire and Smoke Understanding in Multimodal LLMs
[CVPR 2026] SPARROW: Learning Spatial Precision and Temporal Referential Consistency in Pixel-Grounded Video MLLMs
[ICLR 2026] RedSage: A Cybersecurity Generalist LLM. Continued Pretraining and Post-trained RedSage Models.
[ICLR 2026] RedSage: A Cybersecurity Generalist LLM. List of Cybersecurity Benchmarks Datasets.
[ICLR 2026] RedSage: A Cybersecurity Generalist LLM. List of RedSage Datasets
Benchmarks for Evaluating Trade-offs in Image Generation (ICCV 2025, ACMMM 2026)
Visually Grounded Commonsense Reasoning Supervision for CLIP (ECCV 2026)
Visually Grounded Commonsense Reasoning Supervision for CLIP (ECCV 2026)
-
RISys-Lab/ReasonCLIP-B32-S1
Zero-Shot Image Classification • 0.2B • Updated • 4 -
RISys-Lab/ReasonCLIP-L14-224-S1
Zero-Shot Image Classification • 0.4B • Updated • 2 -
RISys-Lab/ReasonCLIP-L14-336-S1
Zero-Shot Image Classification • 0.4B • Updated • 5 -
RISys-Lab/ReasonSigLIP-So14-384-S2
Zero-Shot Image Classification • 0.9B • Updated • 2