#vision (17 Repositories)
Ranked open-source repositories tagged with #vision, scored by pull request acceptance likelihood and maintainer engagement velocity.
23.6%
38.5h
17 repositories tagged #vision
GoogleCloudPlatform/java-docs-samples
Java and Kotlin Code samples used on cloud.google.com
Skyvern-AI/skyvern
Automate browser based workflows with AI
danny-avila/LibreChat
Enhanced ChatGPT Clone: Features Agents, MCP, Skills, DeepSeek, Anthropic, AWS, OpenAI, Responses API, Azure, Groq, o1, GPT-5, Mistral, OpenRouter, Vertex AI, Gemini, Artifacts, AI model switching, message search, Code Interpreter, langchain, DALL-E-3, OpenAPI Actions, Functions, Secure Multi-User Auth, Presets, open-source for self-hosting. Active
vcamapp/app
VTuber Camera, macOS app that shows your avatar using CoreMedia I/O's virtual camera.
ysr666/dsh-vision-router
Eyes for text-only DeepSeek Harness agents: built-in free vision chain (no key) + pixel-level vision tools (Q&A, grounding, crop, pixel diff, colors, OCR, SVG trace, cutout, screenshots). One-command install, no Python, image turns work like ordinary tool-calling turns.
PhotonVision/photonvision
PhotonVision is the free, fast, and easy-to-use computer vision solution for the FIRST Robotics Competition.
xiincs/claude-code-vision-skill
为 Claude Code 赋能多模态视觉能力,支持豆包、通义千问、GPT-4o 等模型,用于截图 / UI / 图表分析;适配 DeepSeek 等无视觉底座,搭配 browser-harness 可做前端布局自动化检查。
StarTrail-org/PixelRAG
https://arxiv.org/abs/2606.28344. The end of web parsing. The beginning of scalable pixel-native search. link: https://pixelrag.ai/
Feghal/ImageDetect
✂️ Detect and crop faces, barcodes and texts in image with iOS 11 Vision api.
andyzeng/arc-robot-vision
MIT-Princeton Vision Toolbox for Robotic Pick-and-Place at the Amazon Robotics Challenge 2017 - Robotic Grasping and One-shot Recognition of Novel Objects with Deep Learning.
liushuangls/go-anthropic
Anthropic Claude API wrapper for Go
OpenMOSS/MOSS-VL
MOSS-VL is the core multimodal model series within the OpenMOSS ecosystem, dedicated to visual understanding.
pmh47/dirt
DIRT: a fast differentiable renderer for TensorFlow
AprilRobotics/apriltag_ros
A ROS wrapper of the AprilTag 3 visual fiducial detector
cocoa-ai/FacesVisionDemo
👀 iOS11 demo application for age and gender classification of facial images.
bytedance/UI-TARS-desktop
The Open-Source Multimodal AI Agent Stack: Connecting Cutting-Edge AI Models and Agent Infra
deltacv/PaperVision
Create your custom OpenCV pipelines using a user-friendly node editor, inspired by industry-leading interfaces! Quickly prototype your vision as you edit.