AI & ML interests
None defined yet.
Recent Activity
Papers
Solvita: Enhancing Large Language Models for Competitive Programming via Agentic Evolution
DR$^{3}$-Eval: Towards Realistic and Reproducible Deep Research Evaluation
models 0
None public yet
datasets 11
NJU-LINK/CoVEBench
Viewer • Updated • 626
NJU-LINK/TELBench
Updated • 26
NJU-LINK/WebCompass
Viewer • Updated • 933 • 22.4k • 6
NJU-LINK/ViDiC-1K
Updated • 402 • 5
NJU-LINK/DR3-Eval
Viewer • Updated • 100 • 2.04k • 2
NJU-LINK/CodeTraceBench
Viewer • Updated • 4.32k • 3.07k • 2
NJU-LINK/OmniVideoBench
Viewer • Updated • 1k • 2.59k • 5
NJU-LINK/camerabench_binary
Viewer • Updated • 7.83k • 23
NJU-LINK/MT-Video-Bench
Updated • 100 • 4
NJU-LINK/T2AV-Compass
Viewer • Updated • 500 • 128 • 4