ai
Pigo is a pure Go face detection, pupil/eyes localization and facial landmark points detection library based on the Pixel Intensity Comparison-based Object detection paper.
Skybox AI uses AI to generate full 360 degree panoramic images. 2 available mode tabs below give you full creative control of your skybox.
Get up and running with large language models, locally. Run Llama 2 and other models on macOS. Customize and create your own.
NASA and IBM have teamed up to create an AI Foundation Model for Earth Observations, using large-scale satellite and remote sensing data, including the Harmonized Landsat and Sentinel-2 (HLS) data. By embracing the principles of open AI and open science, both organizations are actively contributing to the global mission of promoting knowledge sharing and accelerating innovations in addressing critical environmental challenges. With Hugging Face's platform, they simplify geospatial model training and deployment, making it accessible for open science users, startups, and enterprises on multi-cloud AI platforms like watsonx. Additionally, Hugging Face enables easy sharing of the pipelines of the model family, which our team calls Prithvi, within the community, fostering global collaboration and engagement.
The world's simplest facial recognition api for Python and the command line.
Recognize and manipulate faces from Python or from the command line with the world's simplest face recognition library.
Operating LLMs in production.
An open platform for operating large language models (LLMs) in production. Fine-tune, serve, deploy, and monitor any LLMs with ease.
Search 1000s of free seamless HD PBR textures. Create Textures With Poly.
Generate 3D materials with AI in a free online editor, or search our growing community library.
Trace Pixels To Vectors in Full Color, Fully Automatically, Using AI.
Convert your JPEG and PNG bitmaps to SVG vectors quickly and easily. Fully Automatically. Using AI.
Accurate AI Transcriptions in Minutes.
Web service proposing to transcribe video and/or audio content using AI
Free and Open Source Machine Translation API. Self-hosted, offline capable and easy to setup.
Unlike other APIs, it doesn't rely on proprietary providers such as Google or Azure to perform translations. Instead, its translation engine is powered by the open source Argos Translate library.
open-source geolocation.
Yachay is an open-source Machine Learning community. We have collected decades worth of useful natural language data from traditional media (i.e. New York Times articles), social media (i.e. Twitter & Reddit), messenger channels, tech blogs, GitHub profiles and issues, the dark web, and legal proceedings, as well as the decisions and publications of government regulators and legislators all across the world.
Use Kiota to generate API clients to call any OpenAPI-described API.
Kiota is a command line tool for generating an API client to call any OpenAPI described API you are interested in. The goal is to eliminate the need to take a dependency on a different API SDK for every API that you need to call. Kiota API clients provide a strongly typed experience with all the features you expect from a high quality API SDK, but without having to learn a new library for every HTTP API.
Segment Anything Model (SAM): a new AI model from Meta AI that can "cut out" any object, in any image, with a single click.
SAM is a promptable segmentation system with zero-shot generalization to unfamiliar objects and images, without the need for additional training.
Milvus is an open-source vector database built to power embedding similarity search and AI applications. Milvus makes unstructured data search more accessible, and provides a consistent user experience regardless of the deployment environment.
Related contents:
Stability AI Language Models.
This repository contains Stability AI's ongoing development of the StableLM series of language models and will be continuously updated with new checkpoints. The following provides an overview of all currently available models. More coming soon.
Play and create AI-generated adventures with infinite possibilities. Not sure where to start?
K8sGPT is a tool for scanning your kubernetes clusters, diagnosing and triaging issues in simple english. It has SRE experience codified into it’s analyzers and helps to pull out the most relevant information to enrich it with AI.
🎚️ Open Source Audio Matching and Mastering. Matchering 2.0 is a novel Containerized Web Application and Python Library for audio matching and mastering.
It follows a simple idea - you take TWO audio files and feed them into Matchering. Our algorithm matches both of these tracks and provides you the mastered TARGET track with the same RMS, FR, peak amplitude and stereo width as the REFERENCE track has.
A Jasper alternative open source with ChatGPT.
This project uses ChatGPT API to create almost any text based output for your need - from marketing content to blog post ideas and a lot more. It uses simple template based components to ask ChatGPT for generating results Creating new templates or tasks take about 30 mins. no more, so you can extend it for your needs or wait for new template release :)
The free AI encyclopedia. AI tools, podcasts, prompts, newsletter, and movies.
A browser interface based on Gradio library for Stable Diffusion.
OpenChatKit provides a powerful, open-source base to create both specialized and general purpose chatbots for various applications. The kit includes an instruction-tuned 20 billion parameter language model, a 6 billion parameter moderation model, and an extensible retrieval system for including up-to-date responses from custom repositories. It was trained on the OIG-43M training dataset, which was a collaboration between Together, LAION, and Ontocord.ai. Much more than a model release, this is the beginning of an open source project. We are releasing a set of tools and processes for ongoing improvement with community contributions.
Transcribe and translate any audio file.
Free, fast and accurate transcription of audio files. 100% free to use.
Free AI filter for cleaning up spoken audio. Enhance voice recordings for free.
Speech enhancement makes voice recordings sound as if they were recorded in a professional studio.
Generate, Edit & Filter images using the DALL-E 2 API
dallecli is a command line app designed to provide users with the ability to generate, edit and filter images using the DALL-E 2 API provided by OpenAI.
Read less, understand more. Unleash the power of quick and easy reading - just paste your URL for an instant summary!
Welcome to Jotte, an AI-powered graph-based writing tool that helps you create high-quality, informative content with ease. Jotte uses nodes and varying specificities of summaries to carry relevant information through an extremely long set of text, making it an ideal tool for creating long-form content.
Coqui STT (frogSTT) is a fast, open-source, multi-platform, deep-learning toolkit for training and deploying speech-to-text models. frogSTT is battle tested in both production and research rocket
PhotoPrism® is an AI-Powered Photos App for the Decentralized Web.
It makes use of the latest technologies to tag and find pictures automatically without getting in your way. You can run it at home, on a private server, or in the cloud.
The Open Source Privacy-Focused Voice Assistant.
Mycroft is the world’s leading open source voice assistant. It is private by default and completely customizable.
fastai is a deep learning library which provides practitioners with high-level components that can quickly and easily provide state-of-the-art results in standard deep learning domains, and provides researchers with low-level components that can be mixed and matched to build new approaches. It aims to do both things without substantial compromises in ease of use, flexibility, or performance. This is possible thanks to a carefully layered architecture, which expresses common underlying patterns of many deep learning and data processing techniques in terms of decoupled abstractions. These abstractions can be expressed concisely and clearly by leveraging the dynamism of the underlying Python language and the flexibility of the PyTorch library. fastai includes:
Qdrant (read: quadrant ) is a vector similarity search engine and vector database. It provides a production-ready service with a convenient API to store, search, and manage points - vectors with an additional payload. Qdrant is tailored to extended filtering support. It makes it useful for all sorts of neural-network or semantic-based matching, faceted search, and other applications.
Related contents:
An AI-powered Personal Identifiable Information (PII) scanner.. Octopii is an open-source AI-powered Personal Identifiable Information (PII) scanner that can look for image assets such as Government IDs, passports, photos and signatures in a directory.
Free UI faces for designers, avatars, dummy faces, AI generated people faces. UI faces for your web projects and designs AI generated people faces you can use in your works for free
Audio & Video Transcription | Speech-to-text. Smarter subtitling and transcription. We combine artificial and human intelligence to bring you accurate and fast transcripts, captions, and translated subtitles with ease.
Online Photo Editor Photopea.com is a free online tool for editing raster and vector graphics with support for PSD, AI, and Sketch files.
A Stable Diffusion app for macOS built with SwiftUI and Apple's ml-stable-diffusion CoreML models.
Azure Cloud Advocates at Microsoft are pleased to offer a 12-week, 24-lesson curriculum all about Artificial Intelligence.
Meet Zed – The fast, collaborative code editor.
Zed is a high-performance, multiplayer code editor from the creators of Atom and Tree-sitter. It's also open source.
Related contents:
A node-based image processing and AI upscaling GUI that makes it easy to chain together complex processing tasks. A flowchart/node-based image processing GUI aimed at making chaining image processing tasks (especially upscaling done by neural networks) easy, intuitive, and customizable. No existing upscaling GUI gives you the level of customization of your image processing workflow that chaiNNer does. Not only do you have full control over your processing pipeline, you can do incredibly complex tasks just by connecting a few nodes together.
AI for Ansible Content Development. Easily Generate, Customize, & Use! Save time and get unstuck. Tell ansible.ai what you’re thinking to automate in your IT infrastructure and it will generate syntactically correct playbook to help you get there.
Basic Pitch, a free audio-to-MIDI converter with pitch bend detection, built by Spotify. Basic Pitch is a Python library for Automatic Music Transcription (AMT), using lightweight neural network developed by Spotify's Audio Intelligence Lab.
Automate Functional Testing from UI to the API. Our AI-powered functional testing tool accelerates test automation. It works across desktop, web, mobile, mainframe, composite, and packaged enterprise-grade applications.
Using generative adversarial networks (GAN), we can learn how to create realistic-looking fake versions of almost anything, as shown by this collection of sites that have sprung up in the past month.
Remove objects, people, text and defects from any picture for free
An open source platform for the machine learning lifecycle.
MLflow is a platform to streamline machine learning development, including tracking experiments, packaging code into reproducible runs, and sharing and deploying models. MLflow offers a set of lightweight APIs that can be used with any existing machine learning application or library (TensorFlow, PyTorch, XGBoost, etc), wherever you currently run ML code (e.g. in notebooks, standalone applications or the cloud).
Search prompts for Stable Diffusion, DALL-E & Midjourney. Search millions of art images by AI models like DALL-E, Stable Diffusion, Midjourney...
Next-generation creation suite. Everything you need to make content, fast. Magical AI tools, realtime collaboration, precision editing, and more. Your next-generation content creation suite.
Open-Capture is the one and only 100% Open Source intelligent capture managment.
An AI recommendation engine to discover new music, movies, art and more.
an open-source GitHub Copilot server. This is an attempt to build a locally hosted version of GitHub Copilot. It uses the SalesForce CodeGen models inside of NVIDIA's Triton Inference Server with the FasterTransformer backend.
Implementation of paper - YOLOv7: Trainable bag-of-freebies sets new state-of-the-art for real-time object detectors.
DeepFaceLab is the leading software for creating deepfakes.
PaddleNLP is an easy-to-use and powerful natural language processing development library. Aggregates high-quality pre-trained models in the industry and provides an out -of-the-box development experience. The model library covering multiple scenarios of NLP and industrial practice examples can meet the needs of developers for flexible customization .