Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Balamurugan Balakreshnan
Balab2021
5
2
3
Follow
0 followers
·
11 following
https://balakreshnan.github.io/
bala_babal
balakreshnan
balamurugan-balakreshnan
balabala76.bsky.social
AI & ML interests
AI & ML, Deep Learning, Large Language models, Large Vision Models, Large Action Models, Small Language models,
Recent Activity
updated
a model
9 days ago
Balab2021/nemotron-3.5-30b-a3b-lora-finetuned
published
a model
9 days ago
Balab2021/nemotron-3.5-30b-a3b-lora-finetuned
updated
a model
10 days ago
Balab2021/Qwen3-1.7B-GRPO-Math
View all activity
Organizations
Balab2021
's models
48
Sort: Recently updated
Balab2021/nemotron-3.5-30b-a3b-lora-finetuned
Text Generation
•
Updated
9 days ago
•
10
Balab2021/Qwen3-1.7B-GRPO-Math
Updated
10 days ago
Balab2021/Nemotron-3.5-Lightning-30B-A3B-LoRA
Updated
11 days ago
•
38
Balab2021/nemotron-3-nano-finetuned
32B
•
Updated
23 days ago
•
11
Balab2021/llama-3.2-1b-squad-lora
Text Generation
•
Updated
29 days ago
•
12
Balab2021/Qwen3-VL-2B-LLaVA-LoRA
Updated
Jul 29
Balab2021/Qwen2.5-3B-OASST1-QLoRA_h200-run4-higher-capacity
Text Generation
•
Updated
Jul 28
•
2
Balab2021/Qwen2.5-3B-OASST1-QLoRA_h200-run3-low-lr
Text Generation
•
Updated
Jul 28
•
2
Balab2021/Qwen2.5-3B-OASST1-QLoRA_h200-run2-conservative
Text Generation
•
Updated
Jul 28
Balab2021/Qwen2.5-3B-OASST1-QLoRA_h200-run1-balanced
Text Generation
•
Updated
Jul 28
Balab2021/Nemotron-Mini-4B-LoRA-OpenAssistant
Updated
Jul 27
•
1
Balab2021/Qwen2.5-3B-OASST1-QLoRA-run2-conservative
Text Generation
•
Updated
Jul 27
Balab2021/Qwen2.5-3B-OASST1-QLoRA-run1-balanced
Text Generation
•
Updated
Jul 27
Balab2021/Qwen2.5-3B-Instruct-oasst1-lora_160726
Text Generation
•
Updated
Jul 16
Balab2021/qwen-workflow-planner-qwen2p5-lora
Updated
Jun 16
Balab2021/gemma_4_lora
Updated
Apr 17
Balab2021/smol-course-SmolVLM2-2.2B-Instruct-trl-sft-ChartQA
Updated
Oct 11, 2025
Balab2021/gpt-oss-20b-multilingual-reasoner
Updated
Aug 5, 2025
Balab2021/sftqwen_finetuned_model_1-5BHS
Text Generation
•
2B
•
Updated
Jul 28, 2025
•
4
Balab2021/1B_finetuned_llama3.2_HS
Text Generation
•
1B
•
Updated
Jul 25, 2025
•
4
Balab2021/Qwen2-0.5B-GRPO-test
Updated
Jul 11, 2025
Balab2021/ppo-Huggy
Reinforcement Learning
•
Updated
Feb 17, 2025
•
193
Balab2021/Taxi-V3
Reinforcement Learning
•
Updated
Feb 17, 2025
Balab2021/q-FrozenLake-v1-4x4-noSlippery
Reinforcement Learning
•
Updated
Feb 17, 2025
Balab2021/poca-SoccerTwos
Reinforcement Learning
•
Updated
Feb 12, 2025
•
26
Balab2021/dqn-SpaceInvadersNoFrameskip-v4
Reinforcement Learning
•
Updated
Feb 6, 2025
•
1
Balab2021/Reinforce-model-4-2
Reinforcement Learning
•
Updated
Feb 6, 2025
Balab2021/Reinforce-model-4
Reinforcement Learning
•
Updated
Feb 6, 2025
Balab2021/rl_course_vizdoom_health_gathering_supreme
Reinforcement Learning
•
Updated
Feb 5, 2025
Balab2021/ppo-CartPole-v1
Reinforcement Learning
•
Updated
Feb 5, 2025
Previous
1
2
Next