A graph based CoT for math reasoning (DAGGER tokens <<<< CoT Tokens)
Shubhashis Roy Dipta PRO
dipta007
AI & ML interests
Multimodal Understanding, Reasoning, Generation
Organizations
GanitLLM
-
GanitLLM: Difficulty-Aware Bengali Mathematical Reasoning through Curriculum-GRPO
Paper • 2601.06767 • Published -
dipta007/Ganit
Viewer • Updated • 32.3k • 47 -
dipta007/GanitLLM-4B_SFT_CGRPO
Text Generation • 196k • Updated • 1.47k -
dipta007/GanitLLM-4B_SFT_GRPO
Text Generation • 196k • Updated • 4 • 1
VC-Inspector
-
Advancing Reference-free Evaluation of Video Captions with Factual Analysis
Paper • 2509.16538 • Published • 1 -
dipta007/VCInspector-7B
Image-Text-to-Text • 8B • Updated • 13 • 1 -
dipta007/VCInspector-3B
Image-Text-to-Text • 4B • Updated • 17 • 1 -
dipta007/ActivityNet-FG-It
Viewer • Updated • 242k • 36
Q2E: Query-to-Event Decomposition for Zero-Shot Multilingual
Datasets used in the paper: Q2E: Query-to-Event Decomposition for Zero-Shot Multilingual Text-to-Video Retrieval
-
Q2E: Query-to-Event Decomposition for Zero-Shot Multilingual Text-to-Video Retrieval
Paper • 2506.10202 • Published -
dipta007/Q2E_MultiVENT_LLAMA_3.3_70B_InternVL_38B_Funiform_16_noASR
Viewer • Updated • 2.39k • 5 -
dipta007/Q2E_MultiVENT_LLAMA_3.3_70B_InternVL_38B_Funiform_16_ASR
Viewer • Updated • 2.39k • 5 -
dipta007/Q2E_MSRVTT-1kA_LLAMA_3.3_70B_InternVL_38B_Funiform_16_ASR
Viewer • Updated • 1k • 4
DAGGER
A graph based CoT for math reasoning (DAGGER tokens <<<< CoT Tokens)
VC-Inspector
-
Advancing Reference-free Evaluation of Video Captions with Factual Analysis
Paper • 2509.16538 • Published • 1 -
dipta007/VCInspector-7B
Image-Text-to-Text • 8B • Updated • 13 • 1 -
dipta007/VCInspector-3B
Image-Text-to-Text • 4B • Updated • 17 • 1 -
dipta007/ActivityNet-FG-It
Viewer • Updated • 242k • 36
GanitLLM
-
GanitLLM: Difficulty-Aware Bengali Mathematical Reasoning through Curriculum-GRPO
Paper • 2601.06767 • Published -
dipta007/Ganit
Viewer • Updated • 32.3k • 47 -
dipta007/GanitLLM-4B_SFT_CGRPO
Text Generation • 196k • Updated • 1.47k -
dipta007/GanitLLM-4B_SFT_GRPO
Text Generation • 196k • Updated • 4 • 1
Q2E: Query-to-Event Decomposition for Zero-Shot Multilingual
Datasets used in the paper: Q2E: Query-to-Event Decomposition for Zero-Shot Multilingual Text-to-Video Retrieval
-
Q2E: Query-to-Event Decomposition for Zero-Shot Multilingual Text-to-Video Retrieval
Paper • 2506.10202 • Published -
dipta007/Q2E_MultiVENT_LLAMA_3.3_70B_InternVL_38B_Funiform_16_noASR
Viewer • Updated • 2.39k • 5 -
dipta007/Q2E_MultiVENT_LLAMA_3.3_70B_InternVL_38B_Funiform_16_ASR
Viewer • Updated • 2.39k • 5 -
dipta007/Q2E_MSRVTT-1kA_LLAMA_3.3_70B_InternVL_38B_Funiform_16_ASR
Viewer • Updated • 1k • 4