HN user

bleonard

421 karma

CEO and cofounder @ Grouparoo Previously: CTO and technical cofounder @ TaskRabbit

[ my public key: https://keybase.io/bleonard; my proof: https://keybase.io/bleonard/sigs/A_ilc1h8TmtRx3gMAUkpDEdRZIcQ9W-aNqb3yTKQQ3Q ]

Posts30
Comments74
View on HN
news.ycombinator.com 1y ago

Show HN: Airbyte 1.0, Marketplace, AI Assist, GenAI Support and Enterprise GA

bleonard
58pts14
www.opensourcedatastack.com 4y ago

Free / Tomorrow: Open-Source Data Stack Conference

bleonard
2pts0
www.opensourcedatastack.com 4y ago

Open Source Data Stack Conference

bleonard
2pts0
www.grouparoo.com 5y ago

Development Workflow for Reverse ETL

bleonard
18pts1
www.grouparoo.com 5y ago

Data Makes Your Tools Smarter

bleonard
1pts0
www.grouparoo.com 5y ago

Reverse ETL with Dbt and Grouparoo

bleonard
1pts0
www.grouparoo.com 5y ago

Batching API Requests

bleonard
2pts0
www.grouparoo.com 5y ago

Grouparoo: Declarative Data Sync

bleonard
63pts13
www.grouparoo.com 5y ago

Lighthouse Reports as GitHub Comment

bleonard
3pts0
techcrunch.com 5y ago

Grouparoo raises $3M to build open source customer data integration framework

bleonard
3pts0
www.grouparoo.com 5y ago

Sync your data warehouse with Salesforce

bleonard
1pts1
www.grouparoo.com 5y ago

Exports Is Not a Function

bleonard
1pts0
github.com 5y ago

Synchronization Algorithm Exploration

bleonard
1pts0
www.grouparoo.com 6y ago

Simulating Cohorts

bleonard
4pts0
www.inc.com 10y ago

TaskRabbit Quadruples Its Business, Says It Will Turn Profitable in 2016

bleonard
2pts0
tech.taskrabbit.com 10y ago

V2 – Retrospective

bleonard
1pts0
tech.taskrabbit.com 10y ago

React Native Example App: Navigation

bleonard
14pts0
tech.taskrabbit.com 11y ago

Resque Bus renamed and now works with Sidekiq

bleonard
5pts0
nuclide.io 11y ago

Nuclide: An open-source IDE for React Native

bleonard
275pts93
tech.taskrabbit.com 11y ago

Apple TV Dashboard

bleonard
6pts0
tech.taskrabbit.com 11y ago

Translating Rails fields

bleonard
2pts0
www.bleonard.com 11y ago

Rails Stats

bleonard
4pts0
www.bleonard.com 12y ago

iOS Authentication Blocks

bleonard
7pts3
www.cakewalkapp.com 12y ago

Show HN: iPhone app for parents to help play in the real world

bleonard
2pts1
www.bleonard.com 12y ago

Jekyll Slides

bleonard
11pts0
tech.taskrabbit.com 12y ago

Resque Bus

bleonard
56pts15
www.bleonard.com 13y ago

Rails: Rollback ActiveRecord Model in after_save

bleonard
1pts0
tech.taskrabbit.com 13y ago

Rails App Template Alternative

bleonard
12pts4
www.bleonard.com 13y ago

Ruby: Singletons, Threads, and Flexibility

bleonard
58pts20
www.bleonard.com 13y ago

Hubtime: Personal Github Graphs

bleonard
1pts0

A blast from the past. When Taskrabbit was acquired by IKEA, I built several tools that went through the whole catalog via various crawling approaches. One tool was to estimate how long it would be to put each item together for an initial training set.

I am excited about the Rails defaults where background and cache and sockets are all database driven. For normal-sized projects that still need those things, it's a huge win in simplicity.

Ario | Onsite in Palo Alto, CA | Full-time | heyario.com

I recently started at Ario where we are building an app for parents that uses AI to make things easier. We are targeting saving them one hour every day. Previously, I co-founded Taskrabbit. That was somewhat similar but now it's time to get LLMs in on the action! It's in the app store, but still early in the game, so we are looking for engineers to build it out. We are using Python and React Native.

To learn more: brian at heyario.com

Just came to say, I still think this is the best balance between the many factors of running a dev team. I keep trying to recreate it in every tool I use.

Airbyte acquired our Reverse ETL company, Grouparoo, 1.5 years ago. There is so much to solve making just the Extract and Load work well and so much value that comes from that, we have been busy there. I'm excited to circle back to publishing next year.

I like how the article notes that the stuff we were talking about with Reverse ETL (mostly activating your data in SaaS systems like Salesforce, Zendesk, etc) is one important part of Publishing. But we are also seeing traditional use cases like file uploads and new fancy stuff like vector databases.

Maybe some day this will be available in JSON.

I built a system for TaskRabbit that scraped all the IKEA products from a variety of sources and ran algorithms to determine their category and predict how long they would take to be assembled. Then there was a Mechanical Turk sort of system for human input. When combined with real-world feedback from the Taskers, it was pretty good.

For better or worse, I've personally been through the entire catalog multiple times.

There are probably some nuances one level down. Things our users have told us they can do in these areas that, to my knowledge, Hightouch doesn't do:

* Combine data from different sources to define a model. We'v seen using Postgres as a source of truth and supplementing with Snowflake data, for example.

* Add tags to contacts in mailchimp, zendesk or make lists of them in customer.io, Pardot, etc based on segmentation. I believe Hightouch Audiences is more like a filter.

* Full workflow with branches, PRs, test suite in a repo. I saw Hightouch added git syncing to a known branch yesterday and it looks cool, but it's not the full workflow yet.

I'm certainly trying to keep it in the friendly-competition area, especially on this thread :-)

Congrats on the launch! Hightouch looks great and this need is real. Things seem to be going well, so I don't think I'm taking too much away by mentioning that we have been been working on Grouparoo, an open source alternative that solves similar pain points.

A few differences: git developer workflow focused (branches, CI, PRs, etc), ability to self host, segmentation in destinations (tagging people in mailchimp based on rules, for example)

https://www.grouparoo.com

We just finished up the Open Source Data Stack conference, which is all about this topic. Feel free to check out the reply.

Specifically, open source approaches to the modern data stack where the trend is picking the right tools for the job that revolve around the warehouse central data store.

The pieces discussed were around getting data in (Snowplow events, Meltano ELT), transforming it (dbt), reporting (Superset), getting it back into tools (Grouparoo Reverse ETL), and orchestrating things (Dagster).

https://www.opensourcedatastack.com

It's open source and free. We're still figuring out the paid offering. Add yes, it is hard :-)

My experience is that a "contact us" button for a startup is just as likely to be and MVP test as it is to be some nefarious trickery.

We can use any query to bring in and are considering various ways to process it after that. That's in the near term roadmap.

The main kind of "processing" that's done now is using all these properties to calculate cohort membership (High Value users) so that all these tools can use it: Zendesk to route tickets, Marketo to trigger a campaign, even the product to change their dashboard.

There are usual suspects for sources. Databases (MySQL, Postgres) and data warehouse (Snowflake, BigQuery, RedShift).

Though data out is more common, we can also bring in data from any given SaaS tool as well. For example, we have a Mailchimp _source_ that will pull in people as they signup through their form.

There is a plugin model and a few Typescript interfaces to implement to be either a source or a destination.

We were on Hacker News a few months back. Since then, developers have been using the UI to set up automated data movement from their databases to Mailchimp, Marketo, Salesforce, and more.

But we also heard that they wanted it more like their normal development workflow. So we now make it even easier to sync data to cloud-based tools via declarative data models and integrations. With this, you manage data sync just like you would any other part of your stack and Grouparoo takes care of getting the right data to the places you want.

We’re excited and around to answer any questions. - comments here - Slack: https://www.grouparoo.com/chat - Email: brian at grouparoo.com

Just for fun and completeness, there's also taking action on this data outside of your warehouse by making your external tools smarter - in some cases, even writing _back_ to the things you EL'd into it. For example, writing product usage to Salesforce. These are the kinds of things we are focusing on at Grouparoo, also an open source project.

I would recommend separating the service between transactional and marketing. Otherwise, there is always something going on with opt-out.

So use Sendgrid or Mailgun for transactional. Hit their API to send a mail. There are pros/cons around keeping the templates in there, but I would lean that way.

For drip campaigns, I would sync the product database to a tool like Mailchimp or the others ones that have been mentioned. We made an open source way to do that syncing, called Grouparoo.