Gyanateet Dutta
AI & ML interests
Recent Activity
Organizations
-
PhotoVerse: Tuning-Free Image Customization with Text-to-Image Diffusion Models
Paper • 2309.05793 • Published • 51 -
3D Gaussian Splatting for Real-Time Radiance Field Rendering
Paper • 2308.04079 • Published • 203 -
stabilityai/stable-diffusion-xl-base-1.0
Text-to-Image • 3B • Updated • 1.39M • • 7.96k -
Ryukijano/lora-trained-xl-kaggle-p100
Text-to-Image • Updated • 2 • 1
-
Ryukijano/rl_course_vizdoom_health_gathering_supreme
Reinforcement Learning • Updated -
Ryukijano/Mujoco_rl_halfcheetah_Decision_Trasformer
Reinforcement Learning • Updated • 1 -
Ryukijano/poca-SoccerTwos
Reinforcement Learning • Updated • 3 -
AlphaStar Unplugged: Large-Scale Offline Reinforcement Learning
Paper • 2308.03526 • Published • 29
-
NeuroPrompts: An Adaptive Framework to Optimize Prompts for Text-to-Image Generation
Paper • 2311.12229 • Published • 25 - Running on ZeroAgentsFeatured1.02k
IP-Adapter-FaceID
🧑1.02kGenerate AI images featuring your own face
-
Design2Code: How Far Are We From Automating Front-End Engineering?
Paper • 2403.03163 • Published • 98
-
Unsupervised Universal Image Segmentation
Paper • 2312.17243 • Published • 21 -
Denoising Vision Transformers
Paper • 2401.02957 • Published • 32 -
timm/ViT-B-16-SigLIP
Zero-Shot Image Classification • Updated • 112k • 39 - Running on ZeroAgents19
Slimsam
🌖19Small yet powerful mask generation application ⚡️
- PausedAgents68
MeshAnythingV2
🚀68Generate artist-style 3D mesh from your input model
- Runtime errorAgents10
En3D
🏃10 - Running on ZeroAgents57
MASt3R
📉57Create a 3D model from multiple uploaded photos
-
naver/MASt3R_ViTLarge_BaseDecoder_512_catmlpdpt_metric
Image-to-3D • 0.7B • Updated • 27.9k • 20
-
NeRF-Det: Learning Geometry-Aware Volumetric Representation for Multi-View 3D Object Detection
Paper • 2307.14620 • Published • 15 -
LU-NeRF: Scene and Pose Estimation by Synchronizing Local Unposed NeRFs
Paper • 2306.05410 • Published • 4 -
ashawkey/nerf2mesh
Updated • 14 - Build errorFeatured25
NeRF
🔮25
-
LLM in a flash: Efficient Large Language Model Inference with Limited Memory
Paper • 2312.11514 • Published • 264 -
3D-LFM: Lifting Foundation Model
Paper • 2312.11894 • Published • 15 -
SOLAR 10.7B: Scaling Large Language Models with Simple yet Effective Depth Up-Scaling
Paper • 2312.15166 • Published • 62 -
TinyGPT-V: Efficient Multimodal Large Language Model via Small Backbones
Paper • 2312.16862 • Published • 32
-
EVA-GAN: Enhanced Various Audio Generation via Scalable Generative Adversarial Networks
Paper • 2402.00892 • Published • 13 - Running on ZeroMCPFeatured296
MusicGen Streaming
🔥296Generate and stream music from text prompts
- Runtime errorAgents145
Whisper JAX
👀145Transcribe or translate audio from microphone, file, or YouTube
-
Audio Mamba: Bidirectional State Space Model for Audio Representation Learning
Paper • 2406.03344 • Published • 22
- Running on ZeroAgentsFeatured1.2k
Stable Fast 3D
🎮1.2kGenerate a 3D mesh from a single image
- Running on ZeroAgentsFeatured184
Roblox 3D Assets Generator v1
🪄184Create a 3D model from an image in 10 seconds!
- Running on ZeroAgentsFeatured151
LLaMA Mesh
👀151Create 3D mesh by chatting.
-
stabilityai/stable-point-aware-3d
Image-to-3D • 2B • Updated • 2.22k • 352
-
PhotoVerse: Tuning-Free Image Customization with Text-to-Image Diffusion Models
Paper • 2309.05793 • Published • 51 -
3D Gaussian Splatting for Real-Time Radiance Field Rendering
Paper • 2308.04079 • Published • 203 -
stabilityai/stable-diffusion-xl-base-1.0
Text-to-Image • 3B • Updated • 1.39M • • 7.96k -
Ryukijano/lora-trained-xl-kaggle-p100
Text-to-Image • Updated • 2 • 1
-
NeRF-Det: Learning Geometry-Aware Volumetric Representation for Multi-View 3D Object Detection
Paper • 2307.14620 • Published • 15 -
LU-NeRF: Scene and Pose Estimation by Synchronizing Local Unposed NeRFs
Paper • 2306.05410 • Published • 4 -
ashawkey/nerf2mesh
Updated • 14 - Build errorFeatured25
NeRF
🔮25
-
Ryukijano/rl_course_vizdoom_health_gathering_supreme
Reinforcement Learning • Updated -
Ryukijano/Mujoco_rl_halfcheetah_Decision_Trasformer
Reinforcement Learning • Updated • 1 -
Ryukijano/poca-SoccerTwos
Reinforcement Learning • Updated • 3 -
AlphaStar Unplugged: Large-Scale Offline Reinforcement Learning
Paper • 2308.03526 • Published • 29
-
NeuroPrompts: An Adaptive Framework to Optimize Prompts for Text-to-Image Generation
Paper • 2311.12229 • Published • 25 - Running on ZeroAgentsFeatured1.02k
IP-Adapter-FaceID
🧑1.02kGenerate AI images featuring your own face
-
Design2Code: How Far Are We From Automating Front-End Engineering?
Paper • 2403.03163 • Published • 98
-
LLM in a flash: Efficient Large Language Model Inference with Limited Memory
Paper • 2312.11514 • Published • 264 -
3D-LFM: Lifting Foundation Model
Paper • 2312.11894 • Published • 15 -
SOLAR 10.7B: Scaling Large Language Models with Simple yet Effective Depth Up-Scaling
Paper • 2312.15166 • Published • 62 -
TinyGPT-V: Efficient Multimodal Large Language Model via Small Backbones
Paper • 2312.16862 • Published • 32
-
Unsupervised Universal Image Segmentation
Paper • 2312.17243 • Published • 21 -
Denoising Vision Transformers
Paper • 2401.02957 • Published • 32 -
timm/ViT-B-16-SigLIP
Zero-Shot Image Classification • Updated • 112k • 39 - Running on ZeroAgents19
Slimsam
🌖19Small yet powerful mask generation application ⚡️
-
EVA-GAN: Enhanced Various Audio Generation via Scalable Generative Adversarial Networks
Paper • 2402.00892 • Published • 13 - Running on ZeroMCPFeatured296
MusicGen Streaming
🔥296Generate and stream music from text prompts
- Runtime errorAgents145
Whisper JAX
👀145Transcribe or translate audio from microphone, file, or YouTube
-
Audio Mamba: Bidirectional State Space Model for Audio Representation Learning
Paper • 2406.03344 • Published • 22
- Running on ZeroAgentsFeatured1.2k
Stable Fast 3D
🎮1.2kGenerate a 3D mesh from a single image
- Running on ZeroAgentsFeatured184
Roblox 3D Assets Generator v1
🪄184Create a 3D model from an image in 10 seconds!
- Running on ZeroAgentsFeatured151
LLaMA Mesh
👀151Create 3D mesh by chatting.
-
stabilityai/stable-point-aware-3d
Image-to-3D • 2B • Updated • 2.22k • 352
- PausedAgents68
MeshAnythingV2
🚀68Generate artist-style 3D mesh from your input model
- Runtime errorAgents10
En3D
🏃10 - Running on ZeroAgents57
MASt3R
📉57Create a 3D model from multiple uploaded photos
-
naver/MASt3R_ViTLarge_BaseDecoder_512_catmlpdpt_metric
Image-to-3D • 0.7B • Updated • 27.9k • 20