Zero-copy Rust data engine for Python. Memory-safe, blazingly fast, seamless Python integration via Arrow PyCapsule Interface.
-
Updated
May 30, 2026 - Rust
Zero-copy Rust data engine for Python. Memory-safe, blazingly fast, seamless Python integration via Arrow PyCapsule Interface.
rtomde supports building apis quickly based on data warehouse or lakehouse, which converts sql into RESTful easily, commonly as data service, focusing on providing Readable, Testable, Observable, Maintainable Data Engine.
特斯拉自动驾驶架构思想向自驱动实验室 SDL 迁移—— AI for Science 基础设施视角 | Transferring Tesla's autonomous-driving architecture ideas to Self-Driving Labs — an AI for Science infrastructure perspective
A robust In-Memory Database Engine implemented in Java, focusing on professional software architecture, Clean Code, and the application of 7 Design Patterns.
Auto-labeling and VLM scene understanding pipeline for driving data — open-vocabulary detection (Grounding DINO), SAM segmentation, traffic-rule-aware scene tagging (Qwen2.5-VL), pseudo-label mAP evaluation and long-tail mining. Designed to fit in 6 GB VRAM.
Zora is not just another database engine, it's a revolution in data management which is designed to be more. Zora introduces cutting-edge features and a futuristic approach to storing and accessing data.
[German] Code for generating synthetic text images as described in "Synthetic Data for Text Localisation in Natural Images", Ankush Gupta, Andrea Vedaldi, Andrew Zisserman, CVPR 2016.
LOGOS — an event-driven data engine that transforms raw inputs into structured, append-only data within the AIOS system.
🚀 Use your favourite data store like In-Memory cache, MongoDB, Redis and Elastic without worrying about the internal implementations
Automated visual data harvesting and VLM grounding engine for indoor door detection, video frame filtering, pseudo-labeling, and leakage-free YOLO dataset synthesis.
Zyquo Database (ZDB) — a native Apple Silicon data engine in Swift: columnar tables, vectors, documents, mini-SQL, compression, mmap, HTTP dashboard. Zero dependencies.
An immutable database that follows Snowflake architecture, designed for scalable, replayable systems beyond analytics.
Eventual — independent third-party profile of a public API surface, by API Evangelist. Eventual is a data infrastructure company building Daft, an open-source, high-performance data engine for AI and multimodal workloads. Written in Rust with Python and SQL interfaces, Daft lets teams query and process images, audio, video, documents, embeddings, a
Low-latency, multi-threaded data processing system designed to interface directly with low-level hardware
Scale AI — independent third-party profile of a public API surface, by API Evangelist. Scale AI is the data engine for AI. The company turns raw data into training data by combining ML-powered pre-labeling with multi-tier human review, and ships an extensive REST API and SDKs for managing labeling, evaluation, and generative-AI data pipelines.
End-to-end pipeline ingesting product docs from GitHub, refining through bronze/silver/gold on AWS, and serving to an AI agent
Module that supports provisioning an IBM Cloud® Data Engine instance
Local-first Rust data engine: SQL, native structures, lexical/vector search, WAL, MVCC, recovery, and verifiable proofs in one binary.
Curated papers, datasets, systems, and benchmarks for robot data engines across robot-centric, UMI, human/egocentric, and simulation data.
[ICLR 2025] Scalable Benchmarking and Robust Learning for Noise-Free Ego-Motion and 3D Reconstruction from Noisy Video
To associate your repository with the data-engine topic, visit your repo's landing page and select "manage topics."