New and extensible file format for storage of large columnar datasets.
-
Updated
Aug 4, 2026 - C++
New and extensible file format for storage of large columnar datasets.
[데이터 중심 애플리케이션 설계 ] 북 스터디의 Summary Note와 자료 모음집입니다.
A multi-view data system for serving repository context to coding agents.
Benchmark suite for reproducible evaluation of approximate matrix multiplication algorithms.
An implementation of the SIGMOD23 paper: Detect, Distill and Update: Detect, Distill and Update: Learned DB Systems Facing Out of Distribution Data
Technical Product Manager portfolio focused on data systems, product analytics, and scalable backend-driven products
Data-Systems-Project || B+Tree indexing, covering insertion, deletion, range and analysis.B+Tree supports only Integer value.
Semester Project for CSC4210 Computer Architecture, a 4 task processor design.
Framework-neutral C++ implementations of approximate matrix multiplication algorithms with Python bindings.
Build an indentation-sensitive scripting language in Rust with bytecode VM, Python-like syntax, type hints, and fast startup
Data-Systems-Project || simplified relational database management system supporting only integer tables and matrices. Supports two-phase merge sort, buffer management and aggregate queries.
DataSys-maintained VSAG derivative used by CANDOR-Bench for reproducible vector-index evaluation.
Studying how information systems behave under noise, missing context, and bias.
GitHub profile README for Ahmadreza Samadi.
Using PySpark framework(a distributed cluster computing framework) to answer queries from a massive dataset.
Laboratorijske vježbe iz kolegija Operacijski sustavi. Nastajalo tijekom ljetnog semestra 2022.
Scripts and configuration of embedded data acquistion systems at EOL
Add a description, image, and links to the data-systems topic page so that developers can more easily learn about it.
To associate your repository with the data-systems topic, visit your repo's landing page and select "manage topics."