[ICLR2025] LLaVA-HR: High-Resolution Large Language-Vision Assistant
-
Updated
Aug 14, 2024 - Python
[ICLR2025] LLaVA-HR: High-Resolution Large Language-Vision Assistant
Client d'assistant IA pour l'API Albert de la DiNum
A highly realistic Telegram AI userbot powered by Google Gemini. It acts as your digital clone, featuring human-like typing, media analysis, and a secure control panel.
Enterprise AI assistant automating IT/HR support via customizable Graph-RAG pipelines, real-time tool calling, and document parsing. Engineered for high-security environments using 2FA, personalized dashboards, and local deployment. Reduces internal workloads with sub 3 second retrieval times.
RideGuide Flutter App: A Multimodal Conversational Interface for Passenger-Engaged Spatial Learning in Robotaxis
Clara: An agentic multimodal AI assistant that can see through your webcam, listen to your voice, think with Gemini, and speak back using ElevenLabs. Built with LangGraph, OpenCV, Groq, and Gradio.
A Multimodal AI assistant for automated plant pathology diagnosis, featuring an interactive UI built with Gradio. Developed during my MSc AI at BSBI, the system utilizes a fine-tuned EfficientNet-B0 architecture for high-accuracy image classification, integrated with natural language processing to provide real-time disease identification.
To associate your repository with the multimodal-chatbot topic, visit your repo's landing page and select "manage topics."