en
Feedback
Artificial Intelligence l l AI Updates

Artificial Intelligence l l AI Updates

Open in Telegram
1 617
Subscribers
-324 hours
-17 days
No data30 days
Posts Archive
πŸ”— GitHub_Link ❇️ RealmDreamer: Text-Driven 3D Scene Generation with Inpainting and Depth Diffusion #3D Join my channel: πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡ https://t.me/Artificial_Intelligence_Updates

πŸ”— GitHub_Link ❇️ VSP-LLM (Visual Speech Processing incorporated with LLMs) #LLMs Join my channel: πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡ https://t.me/Artificial_Intelligence_Updates

πŸ”— GitHub_Link ❇️ In-Context Edit: Enabling Instructional Image Editing with In-Context Generation in Large Scale Diffusion Transformer πŸ”₯ Join my channel: πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡ https://t.me/Artificial_Intelligence_Updates

πŸ”— GitHub_Link ❇️ ML-powered speech recognition directly in your browser #STT Join my channel: πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡ https://t.me/Artificial_Intelligence_Updates

πŸ”— GitHub_Link ❇️ A react frontend for segmentation and subject extraction using Cartoon Segmentation πŸ”₯ Join my channel: πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡ https://t.me/Artificial_Intelligence_Updates

πŸ”— GitHub_Link ❇️ FastVLM: Efficient Vision Encoding for Vision Language Models #VLMs Join my channel: πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡ https://t.me/Artificial_Intelligence_Updates

πŸ”— GitHub_Link ❇️ TransPixeler: Advancing Text-to-Video Generation with Transparency πŸ”₯ #txt2vid #img2vid Join my channel: πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡ https://t.me/Artificial_Intelligence_Updates

πŸ”— GitHub_Link ❇️πŸ”₯ Fine-tuning Qwen2.5-VL-3BπŸ”₯ #LLMs Join my channel: πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡ https://t.me/Artificial_Intelligence_Upda
πŸ”— GitHub_Link ❇️πŸ”₯ Fine-tuning Qwen2.5-VL-3BπŸ”₯ #LLMs Join my channel: πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡ https://t.me/Artificial_Intelligence_Updates

πŸ”— GitHub_Link ❇️ HiFi-Score: Fine-grained Image Description Evaluation with Hierarchical Parsing Graphs πŸ”₯ Join my channel:
πŸ”— GitHub_Link ❇️ HiFi-Score: Fine-grained Image Description Evaluation with Hierarchical Parsing Graphs πŸ”₯ Join my channel: πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡ https://t.me/Artificial_Intelligence_Updates

πŸ”— GitHub_Link ❇️ Count Anything πŸ”₯πŸ”₯πŸ”₯ #SAM Join my channel: πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡ https://t.me/Artificial_Intelligence_Updates
πŸ”— GitHub_Link ❇️ Count Anything πŸ”₯πŸ”₯πŸ”₯ #SAM Join my channel: πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡ https://t.me/Artificial_Intelligence_Updates

πŸ”— GitHub_Link ❇️ CFG-Zero*: Improved Classifier-Free Guidance for Flow Matching Models πŸ”₯ Join my channel: πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡ https://t.me/Artificial_Intelligence_Updates

πŸ”— GitHub_Link ❇️ MLflow for Machine Learning Development #Tutorial Join my channel: πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡ https://t.me/Artificial_Int
πŸ”— GitHub_Link ❇️ MLflow for Machine Learning Development #Tutorial Join my channel: πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡ https://t.me/Artificial_Intelligence_Updates

πŸ”— GitHub_Link ❇️ CoMotion: Concurrent Multi-person 3D Motion #3D Join my channel: πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡ https://t.me/Artificial_Intelligence_Updates

πŸ”— GitHub_Link ❇️ Event-Enhanced Blurry Video Super-Resolution πŸ”₯πŸ”₯πŸ”₯ Join my channel: πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡ https://t.me/Artificial_Intelligence_Updates

πŸ”— GitHub_Link ❇️ AirCanvas: it is is a computer vision project that lets you draw using hand gestures. #Tutorial Join my channel: πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡ https://t.me/Artificial_Intelligence_Updates

πŸ“Œ Sources Advances and Challenges in Foundation Agents https://arxiv.org/pdf/2504.01990v1 Why do Multi-agent systems fail ht
πŸ“Œ Sources Advances and Challenges in Foundation Agents https://arxiv.org/pdf/2504.01990v1 Why do Multi-agent systems fail https://arxiv.org/pdf/2503.13657 API vs. GUI Agents - Divergence and Convergence https://arxiv.org/html/2503.11069v1 PaperBench - Evaluating AI's Ability to Replicate Research https://arxiv.org/abs/2504.01848 MemInsight: Autonomous Memory Augmentation for LLM Agents https://arxiv.org/pdf/2503.21760 BEARCUBS: A Benchmark for Computer-Using Web Agents https://arxiv.org/pdf/2503.07919 AgentRxiv: Towards Collaborative Autonomous Research https://arxiv.org/abs/2503.18102 PLAY2PROMPT: Zero-shot Tool Instruction Optimization for LLM Agents https://arxiv.org/abs/2503.14432 Agents Play Thousands of 3D Video Games using PORTAL https://arxiv.org/abs/2503.13356 Join my channel: πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡ https://t.me/Artificial_Intelligence_Updates

πŸ”— GitHub_Link ❇️ DSO: Aligning 3D Generators with Simulation Feedback for Physical Soundness #3D Join my channel: πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡ https://t.me/Artificial_Intelligence_Updates

πŸ”— GitHub_Link ❇️ Geo4D: Leveraging Video Generators for Geometric 4D Scene Reconstruction #4D Join my channel: πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡ https://t.me/Artificial_Intelligence_Updates

πŸ”— GitHub_Link ❇️ Packing Input Frame Context in Next-Frame Prediction Models for Video Generation πŸ”₯πŸ”₯πŸ”₯ Join my channel: πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡ https://t.me/Artificial_Intelligence_Updates

πŸ”— GitHub_Link ❇️ Pusa: Thousands Timesteps Video Diffusion Model #txt2img #img2vid Join my channel: πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡πŸ‘‡ https://t.me/Artificial_Intelligence_Updates