HN user

canadi

66 karma
Posts13
Comments9
View on HN

add_docs() API always UPSERTS and so yes, updates are through "_id". The system auto-assigns an "_id" only when it is not supplied by the user or when an existing field is not mapped as the "_id" field at collection creation time. You will have to use delete_docs() before add_docs(), if you want replace-document behavior.

Our backend architecture is quite scalable and actually grows and shrinks with the demand continuously.

And yes, all documents are automatically indexed and replicated for fast query performance, which is more expensive than just storing them in "_id"->"doc" format. For our use cases and value prop, this one time indexing cost pays for itself several times over by saving time during query processing.

Hey, Igor from Rockset here.

Rockset’s primary use-cases are: 1/ developers building low-latency operational applications, esp. combining real-time data sets with other structured data sets (eg: you are building a microservice to relieve pressure from your OLTP system) 2/ data scientists wanting to quickly test hypotheses on different structured and semi-structured datasets without having to stand-up any servers or do any ETL or data prep. (you can suspend collections/documents in Rockset when you don’t use them -- our pricing page currently only lists Active Documents’ pricing)

Rockset is mutable which allows it to keep itself in sync with any data source, unlike columnar data warehouses, which are not optimized for data manipulation.

Rockset’s strong dynamic typing allows it to treat JSON as a data representation format rather than a special data type or a storage format. So, once you load JSON data into Rockset, you can access all fields at all levels without any special JSON operators or functions.

Comparing Snowflake with Rockset is perhaps akin to comparing Teradata with Elasticsearch. Both useful systems but built for very different use cases.

The biggest thing Rockset has in common with Snowflake is in sharing the philosophy that data management systems have to be built ground up for the cloud to take full advantage of cloud economics. Our blog (https://rockset.com/blog/) has a few posts on these already and we will write more.

Rockset | Senior Software Engineer, Lead Front-end Engineer | San Mateo, CA | Onsite | Full time

At Rockset we are building the next generation of cloud-native data infrastructure. Our team includes founding members of RocksDB, Hadoop Distributed File System, Facebook's search engine (Unicorn) and social graph serving engine (TAO). We are backed by Greylock Partners and Sequoia Capital. We are building our infrastructure on top of Kubernetes on AWS, and are using systems like RocksDB, Kafka, Zookeeper, gRPC and Terraform. Most of our codebase is in C++ and Java.

Open Roles: https://rockset.com/careers (also links to a page where you can apply)

Rockset | Senior Software Engineer | San Mateo, CA | Onsite | Full time

At Rockset we are building the next generation of cloud-native data infrastructure. Our team includes founding members of RocksDB, Hadoop Distributed File System, Facebook's search engine (Unicorn) and social graph serving engine (TAO). We are backed by Greylock Partners and Sequoia Capital. We are building our infrastructure on top of Kubernetes on AWS, and are using systems like RocksDB, Kafka, Zookeeper, gRPC and Terraform. Most of our codebase is in C++ and Java.

Open Roles: https://rockset.com/careers (also links to a page where you can apply)

Rockset | Senior Infastructure Engineer, Lead Frontend Engineer, Software Engineer | San Mateo, CA | Onsite | Full time At Rockset we are building the next generation of cloud-native data infrastructure. Our team includes founding members of RocksDB, Hadoop Distributed File System, Facebook's search engine (Unicorn) and social graph serving engine (TAO). We are backed by Greylock Partners and Sequoia Capital.

We are building our infrastructure on top of Kubernetes on AWS, and are using systems like RocksDB, Kafka, Zookeeper, gRPC and Terraform. Most of our codebase is in C++ and Java.

Open Roles: https://rockset.com/careers (also links to a page where you can apply)

Rockset | Senior Infastructure Engineer, Lead Frontend Engineer, Software Engineer | San Mateo, CA | Onsite | Full time

At Rockset we are building the next generation of cloud-native data infrastructure. Our team includes founding members of RocksDB, Hadoop Distributed File System, Facebook's search engine (Unicorn) and social graph serving engine (TAO). We are backed by Greylock Partners and Sequoia Capital.

We are building our infrastructure on top of Kubernetes on AWS, and are using systems like RocksDB, Kafka, Zookeeper, gRPC and Terraform. Most of our codebase is in C++ and Java.

Open Roles: https://rockset.io/careers To apply, email us at jobs@rockset.io

Rockset | Senior Infastructure Engineer, Lead Frontend Engineer, Software Engineer | San Mateo, CA | Onsite | Full time

At Rockset we are building the next generation of cloud-native data infrastructure. Our team includes founding members of RocksDB, Hadoop Distributed File System, Facebook's search engine (Unicorn) and social graph serving engine (TAO). We are backed by Greylock Partners and Sequoia Capital.

We are building our infrastructure on top of Kubernetes on AWS, and are using systems like RocksDB, Kafka, Zookeeper, gRPC and Terraform. Most of our codebase is in C++ and Java.

Open Roles: https://rockset.io/careers

To apply, email us at jobs@rockset.io