#Gpt-4-vision
Showing 9 of 9 repositories tagged #gpt-4-vision, ranked by stars
Resource, examples & tutorials for multimodal AI, RAG and agents using vector search and LLMs
The most advanced Web UI for AI chat
【新增智能体模式】安卓端全场景GPT助手,可用音量键唤起并进行语音交流,支持联网、拍照、模板、附件解析、智能体模式等 | GPT assistant for Android, activated via volume keys for voice interaction, supporting features such as networking, taking photos, templates, parsing PDF and Office documents, and agent mode.
Cool experiments at the intersection of Computer Vision and Sports ⚽🏃
SGPT is a command-line tool that provides a convenient way to interact with OpenAI models, enabling users to run queries, generate shell commands and produce code directly from the terminal.
Convert a screenshot to a working Flutter app.
A versatile multi-modal chat application that enables users to develop custom agents, create images, leverage visual recognition, and engage in voice interactions. It integrates seamlessly with local LLMs and commercial models like OpenAI, Gemini, Perplexity, and Claude, and allows to converse with uploaded documents and websites.
Extract information, summarize, ask questions, and search videos using OpenAI's Vision API 🚀🎦
ChatGPT wrapper in your TTY