multimodel
Here are 84 public repositories matching this topic...
DeepResearchAgent is a hierarchical multi-agent system designed not only for deep research tasks but also for general-purpose task solving. The framework leverages a top-level planning agent to coordinate multiple specialized lower-level agents, enabling automated task decomposition and efficient execution across diverse and complex domains.
-
Updated
May 4, 2026 - Python
RMDL: Random Multimodel Deep Learning for Classification
-
Updated
Apr 22, 2026 - Python
Awesome_Multimodel is a curated GitHub repository that provides a comprehensive collection of resources for Multimodal Large Language Models (MLLM). It covers datasets, tuning techniques, in-context learning, visual reasoning, foundational models, and more. Stay updated with the latest advancement.
-
Updated
Jul 3, 2026
yolov3, yolo12, dino, segmenations, face, pose, keypoints on deepstream
-
Updated
Dec 7, 2025 - Jupyter Notebook
🧘🏻♂️KarmaVLM (相生):A family of high efficiency and powerful visual language model.
-
Updated
Apr 29, 2024 - Python
This is our solution for KDD Cup 2020. We implemented a very neat and simple neural ranking model based on siamese BERT which ranked first among the solo teams and ranked 12th among all teams on the final leaderboard.
-
Updated
Jun 20, 2020 - Jupyter Notebook
OpenVINO+NCS2/NCS+MutiModel(FaceDetection, EmotionRecognition)+MultiStick+MultiProcess+MultiThread+USB Camera/PiCamera. RaspberryPi 3 compatible. Async.
-
Updated
Feb 12, 2023 - Python
Open MobileUI is a production-ready, cross-platform mobile application that brings the full power of Open WebUI to iOS and Android devices. Built with React Native by zha0090, Open MobileUI transforms your self-hosted AI assistant into a native mobile experience.
-
Updated
May 23, 2026 - TypeScript
End-to-End AI Voice Assistant pipeline with Whisper for Speech-to-Text, Hugging Face LLM for response generation, and Edge-TTS for Text-to-Speech. Features include Voice Activity Detection (VAD), tunable parameters for pitch, gender, and speed, and real-time response with latency optimization.
-
Updated
Apr 15, 2026 - Jupyter Notebook
An open research platform for wearable sensing, nearby-device local multimodal inference, and local-first visual assistance.
-
Updated
Aug 20, 2026 - Python
Accepted by TMM 2022
-
Updated
Aug 18, 2022 - Python
LLM Council works together to answer your hardest questions
-
Updated
Dec 3, 2025 - Python
ArangoGraph is the easiest way to run ArangoDB. Available on AWS and Google Cloud.
-
Updated
Feb 26, 2024
The unified, OpenAI-compatible AI router. 12+ providers, 100+ models, automatic fallback. Built solo from a rented room in Peru. Star + donate to keep it alive.
-
Updated
Aug 23, 2026 - Python
开源 Android 手机自动化 Agent:用自然语言驱动多模态大模型理解屏幕、规划任务并跨应用执行操作,支持可复用 Skills。
-
Updated
Aug 4, 2026 - Kotlin
Robust particle filter based on dynamic averaging of multiple noise models
-
Updated
Nov 14, 2019 - MATLAB
This project is a multi-modal model that works with multiple models combined and accepts audio, images, and text as inputs, generating corresponding audio, images, and text outputs.
-
Updated
Feb 26, 2024 - Python
Add this topic to your repo
To associate your repository with the multimodel topic, visit your repo's landing page and select "manage topics."