Webinar-Lottie.svg

lakeFS Acquires DVC, Uniting Data Version Control Pioneers to Accelerate AI-Ready Data

webcros.svg

Webinar October 22: Governing Autonomous AI Agents with lakeFS + MinIO

The lakeFS Blog

Learn how to use lakeFS

The lakeFS Katacoda Sandbox Environment – Interactive Data Versioning Learning

If you’re interested in playing around and exploring lakeFS, you can now easily get started using the Katacoda demo which provides a personalized sandboxed environment

lakeview - Visibility Tool for AWS S3 Based Data Lakes

Introducing lakeview: A Visibility Tool for AWS S3 Based Data Lakes

Lakeview is a new open source visibility tool for AWS S3 based data lakes. Think of it as ncdu, but for Petabyte-scale data. It’s goal

How to Manage Your Data the Way You Manage Your Code

How to Manage Your Data the Way You Manage Your Code

50 years ago it was very hard to collaborate over code. When developing large scale software projects it was difficult to manage changes to source

Improving Postgres performance tenfold using Go Concurrency

Improving Postgres Performance Tenfold Using Go Concurrency

In this article I will show how Go concurrency enabled us to cut through a daunting DB performance barrier. This blog post continues our journey

Caching in Go

In-process Caching In Go: Scaling lakeFS to 100k Requests/Second

This is a first in a series of posts describing our journey of scaling lakeFS. In this post we describe how adding an in-process cache

Diary of a Data Engineer

Diary of a Data Engineer

A glimpse into the life of a data engineer. Day 1: Finally, an easy one Got a pretty simple task for a change – read

How to pick the right Postgres

How to Pick the Right Postgres for your Application

Lots of applications require a Postgres database. Before you can install them, you will need a Postgres database. How do you pick the right Postgres

Running Presto Locally

The Quick Guide for Running Presto Locally on S3

This post aims to cover our experience running Presto in a local environment with the ability to query Amazon S3 and other S3 Compatible Systems.

Hive Metastore vs AWS Glue

Metadata Management: Hive Metastore vs AWS Glue

Introducing the concept of metadata catalogs and explaining the benefits and pains of using Hive Metastore and AWS Glue
Getting Started

From Zero to Versioned Data in Spark

This tutorial aims to give you a fast start with lakeFS and use its git-like terminology in Spark. It covers the following: This simple flow

lakefs is live

Why We Built lakeFS: Atomic and Versioned Data Lake Operations

lakeFS is an open source platform that delivers resilience and manageability to your existing object-storage based data lake. With lakeFS you can build repeatable, atomic
[hubspot type=form portal=8040338 id=9f5646ec-3e20-4568-9d6e-b82fca022065]