awesome-repositories.com
Blog
MCP
awesome-repositories.com

Descubre los mejores repositorios open-source con nuestra búsqueda potenciada por IA.

ExplorarBúsquedas curadasAlternativas open-sourceSoftware autohospedableBlogMapa del sitio
ProyectoAcerca deCómo clasificamosPrensaServidor MCP
Aviso legalPrivacidadTérminos
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
Back to intake/intake

Open-source alternatives to Intake

30 open-source projects similar to intake/intake, ranked by how many features they have in common. Compare stars, activity and what each one does to find the best Intake alternative.

  • huggingface/datasetsAvatar de huggingface

    huggingface/datasets

    21,643Ver en GitHub↗

    Datasets is a library designed for the management, processing, and sharing of large-scale data collections for machine learning workflows. It functions as both a data processing framework and a versioning platform, providing tools to organize, filter, and transform massive datasets while ensuring reproducibility across research and development teams. The library distinguishes itself by enabling the handling of datasets that exceed available system memory. It utilizes memory-mapped file access, disk-based caching, and lazy iterative streaming to maintain performance when working with large-sca

    Pythonaiartificial-intelligencecomputer-vision
    Ver en GitHub↗21,643
  • activeloopai/hubAvatar de activeloopai

    activeloopai/Hub

    9,177Ver en GitHub↗

    Hub is a multimodal AI data lake and vector database designed for storing and querying embeddings, text, audio, and images. It functions as a dataset version control system and a machine learning data streaming engine to support large-scale model training. The system utilizes a serverless PostgreSQL vector store to index high-dimensional embeddings for semantic search. It provides a visual interface for inspecting multimodal datasets and viewing annotations such as bounding boxes and masks. The platform handles cloud-agnostic storage synchronization and implements lazy, compressed data strea

    C++
    Ver en GitHub↗9,177
  • attaswift/btreeAvatar de attaswift

    attaswift/BTree

    1,324Ver en GitHub↗

    Fast sorted collections for Swift using in-memory B-trees

    Swift
    Ver en GitHub↗1,324
  • ahmed-ali/jsonexportAvatar de Ahmed-Ali

    Ahmed-Ali/JSONExport

    4,812Ver en GitHub↗

    JSONExport is a multi-language code generator and JSON schema converter that transforms JSON data structures into strongly typed source code classes. It serves as an API response mapper, converting JSON objects into data transfer objects to automate the creation of model classes. The tool specializes in multi-language model synthesis, allowing users to define data models across different programming languages using a single JSON input. It generates class boilerplate, including constructors and accessors, and provides a preview pipeline to review the resulting source code before it is saved.

    Swift
    Ver en GitHub↗4,812

Búsqueda con IA

Explora más repositorios increíbles

Describe lo que necesitas en lenguaje sencillo: la IA clasifica miles de proyectos open-source curados por relevancia.

Find more with AI search
  • ahupp/python-magicAvatar de ahupp

    ahupp/python-magic

    2,886Ver en GitHub↗

    python-magic is a C-binding wrapper that provides a Python interface for the libmagic system library. It functions as a file signature analyzer and MIME type detector, identifying file formats by comparing header bytes against a database of known binary signatures. The library enables the identification of file types from both file paths and raw data buffers. It supports custom file signature matching through the injection of user-provided magic databases, allowing for the detection of specialized or proprietary formats. The project covers binary data analysis and MIME type mapping to transl

    Python
    Ver en GitHub↗2,886
  • algolia/algoliasearch-railsAvatar de algolia

    algolia/algoliasearch-rails

    420Ver en GitHub↗

    AlgoliaSearch integration to your favorite ORM

    Ruby
    Ver en GitHub↗420
  • amuste/dnetindexeddbAvatar de amuste

    amuste/DnetIndexedDb

    107Ver en GitHub↗

    Blazor Library for IndexedDB DOM API

    JavaScript
    Ver en GitHub↗107
  • ankane/groupdateAvatar de ankane

    ankane/groupdate

    3,888Ver en GitHub↗

    Groupdate is a PostgreSQL time series aggregator and date grouping tool. It provides a set of SQL functions to group and aggregate temporal records into discrete buckets, such as days, weeks, or months, to calculate sums and averages for reports. The project focuses on ensuring continuous timelines through time series gap filling, which inserts default values for periods where no data exists. It also includes a temporal data formatter that converts grouped date-time keys into localized strings or custom formatting patterns. The tool covers broad temporal data operations, including time range

    Ruby
    Ver en GitHub↗3,888
  • ankane/lockboxAvatar de ankane

    ankane/lockbox

    1,598Ver en GitHub↗

    Modern encryption for Ruby and Rails

    Ruby
    Ver en GitHub↗1,598
  • ankane/rollupAvatar de ankane

    ankane/rollup

    350Ver en GitHub↗

    Rollup time-series data in Rails

    Ruby
    Ver en GitHub↗350
  • ankane/searchkickAvatar de ankane

    ankane/searchkick

    6,717Ver en GitHub↗

    Searchkick is an integration library and wrapper that connects application models to search engines such as Elasticsearch and OpenSearch. It functions as a search index synchronizer, automatically mirroring database records to a search server to enable full-text and vector retrieval. The project provides a high-level interface for implementing keyword search, semantic vector search, and hybrid search. It distinguishes itself through the ability to combine traditional keyword matching with vector embeddings using reranking and fusion techniques to improve precision. The library covers the end

    Ruby
    Ver en GitHub↗6,717
  • ankane/troveAvatar de ankane

    ankane/trove

    80Ver en GitHub↗

    Deploy machine learning models in Ruby (and Rails)

    Ruby
    Ver en GitHub↗80
  • appliedtrust/traildashAvatar de AppliedTrust

    AppliedTrust/traildash

    358Ver en GitHub↗

    AWS CloudTrail Dashboard

    Go
    Ver en GitHub↗358
  • aptabase/aptabaseAvatar de aptabase

    aptabase/aptabase

    1,635Ver en GitHub↗
    TypeScriptanalyticsandroidelectron
    Ver en GitHub↗1,635
  • addresscloud/aws-lambda-docker-rasterioAvatar de addresscloud

    addresscloud/aws-lambda-docker-rasterio

    19Ver en GitHub↗

    AWS Lambda Container Image with Python Rasterio for querying Cloud Optimised GeoTiffs.

    Python
    Ver en GitHub↗19
  • aws/amazon-cognito-androidAvatar de aws

    aws/amazon-cognito-android

    31Ver en GitHub↗

    ARCHIVED: Use https://github.com/aws/aws-sdk-android/

    Java
    Ver en GitHub↗31
  • aws/amazon-cognito-dotnetAvatar de aws

    aws/amazon-cognito-dotnet

    10Ver en GitHub↗

    Official repository for Amazon Cognito Sync Manager SDK for Dotnet.

    C#
    Ver en GitHub↗10
  • aws/amazon-cognito-iosAvatar de aws

    aws/amazon-cognito-ios

    33Ver en GitHub↗

    ARCHIVED: Use https://github.com/aws/aws-sdk-ios/

    Objective-C
    Ver en GitHub↗33
  • aws/amazon-cognito-jsAvatar de aws

    aws/amazon-cognito-js

    199Ver en GitHub↗

    Amazon Cognito Sync Manager for JavaScript

    JavaScript
    Ver en GitHub↗199
  • aws/aws-cloudtrail-processing-libraryAvatar de aws

    aws/aws-cloudtrail-processing-library

    95Ver en GitHub↗

    The AWS CloudTrail Processing Library helps Java developers to easily consume and process log files from AWS CloudTrail.

    Java
    Ver en GitHub↗95
  • aws/aws-dotnet-session-providerAvatar de aws

    aws/aws-dotnet-session-provider

    43Ver en GitHub↗

    A session state provider for ASP.NET applications that stores the sessions in Amazon DynamoDB

    C#
    Ver en GitHub↗43
  • aws/aws-dotnet-trace-listenerAvatar de aws

    aws/aws-dotnet-trace-listener

    15Ver en GitHub↗

    A trace listener for System.Diagnostics that can be used to log events straight to Amazon DynamoDB.

    C#
    Ver en GitHub↗15
  • aws/aws-dynamodb-session-tomcatAvatar de aws

    aws/aws-dynamodb-session-tomcat

    98Ver en GitHub↗

    ARCHIVED: Amazon DynamoDB based session store for Apache Tomcat

    Java
    Ver en GitHub↗98
  • aws/aws-sessionstore-dynamodb-rubyAvatar de aws

    aws/aws-sessionstore-dynamodb-ruby

    68Ver en GitHub↗

    Handles sessions for Ruby web applications using DynamoDB as a backend.

    Ruby
    Ver en GitHub↗68
  • awslabs/amazon-appstream-netA

    awslabs/amazon-appstream-net

    0Ver en GitHub↗
    Ver en GitHub↗0
  • awslabs/amazon-appstream-sample-entitlement-serviceA

    awslabs/amazon-appstream-sample-entitlement-service

    0Ver en GitHub↗
    Ver en GitHub↗0
  • awslabs/amazon-cognito-developer-authentication-sampleAvatar de awslabs

    awslabs/amazon-cognito-developer-authentication-sample

    100Ver en GitHub↗

    Overview

    Java
    Ver en GitHub↗100
  • awslabs/amazon-cognito-streams-sampleAvatar de awslabs

    awslabs/amazon-cognito-streams-sample

    10Ver en GitHub↗

    Sample demonstrating consuming Amazon Cognito Streams

    Java
    Ver en GitHub↗10
  • awslabs/api-gateway-secure-pet-storeAvatar de awslabs

    awslabs/api-gateway-secure-pet-store

    307Ver en GitHub↗

    Amazon API Gateway sample using Amazon Cognito credentials through AWS Lambda

    Objective-C
    Ver en GitHub↗307
  • activeloopai/deeplakeAvatar de activeloopai

    activeloopai/deeplake

    9,175Ver en GitHub↗

    DeepLake is AI data infrastructure consisting of a multimodal data lake, a hybrid search engine, and a serverless vector database. It provides a PostgreSQL-based AI data runtime that combines multimodal storage with streaming pipelines to load and shuffle datasets from cloud storage directly into deep learning training pipelines. The system utilizes lazy indexing to store and slice images, audio, and video without loading entire files into memory. It enables retrieval-augmented generation by persisting high-dimensional embeddings in a serverless vector store and implementing hybrid search tha

    C++agentagentic-ragai
    Ver en GitHub↗9,175