HN user

LiveTheDream

7,980 karma

twitter.com/tobym

software developer, architect, leader, and educator

scala, python, go

hn @ mail.pritama.com

Posts579
Comments499
View on HN
news.mit.edu 7mo ago

Enabling small language models to solve complex reasoning tasks

LiveTheDream
7pts0
www.helpnetsecurity.com 1y ago

Apple plugs zero-day holes used in targeted iPhone attacks

LiveTheDream
2pts0
simonwillison.net 1y ago

Ask questions of SQLite databases and CSV/JSON files in your terminal

LiveTheDream
2pts0
blog.ometer.com 4y ago

Callbacks, Synchronous and Asynchronous (2011)

LiveTheDream
28pts3
www.washingtonpost.com 4y ago

Pandora Papers reveal secret offshore financial system for global elites

LiveTheDream
3pts0
github.com 7y ago

GitHub will render out-of-tree commits

LiveTheDream
16pts3
caddyserver.com 8y ago

Detecting HTTPS Interception – Caddy

LiveTheDream
19pts2
www.wired.com 8y ago

Meet Alex, the Russian Casino Hacker Who Makes Millions Targeting Slot Machines

LiveTheDream
4pts0
arxiv.org 10y ago

Master of Puppets: Analyzing and Attacking a Botnet for Fun and Profit

LiveTheDream
32pts1
sheriff.dynu.com 10y ago

Sheriff – Detecting Price Discrimination

LiveTheDream
10pts0
tech.gc.com 11y ago

These Six Shocking Facts About Meetings Will Change Your Life

LiveTheDream
2pts0
blog.higher-order.com 11y ago

Easy Performance Wins with Scalaz

LiveTheDream
1pts0
groups.google.com 11y ago

JMH vs. Caliper: reference thread

LiveTheDream
8pts0
www.eff.org 11y ago

Lenovo is breaking HTTPS security on its recent laptops

LiveTheDream
20pts1
developerblog.redhat.com 11y ago

Microservice Principles and Immutability – Demonstrated with Spark and Cassandra

LiveTheDream
2pts0
drill.apache.org 11y ago

Apache Drill Graduates to a Top-Level Project

LiveTheDream
2pts0
github.com 11y ago

Emacs advanced Kit focused on Evil-mode

LiveTheDream
4pts0
blog.confluent.io 11y ago

Announcing Confluent, a Company for Apache Kafka and Realtime Data

LiveTheDream
10pts1
kellabyte.com 11y ago

The 99th percentile matters

LiveTheDream
2pts1
openstandard.mozilla.org 11y ago

The Race to Replace the Mythical Anonabox

LiveTheDream
4pts0
latencytipoftheday.blogspot.com 11y ago

Most page loads will experience the 99%'ile server response

LiveTheDream
3pts0
www.hakkalabs.co 11y ago

Data Pipeline at Tapad

LiveTheDream
3pts0
eclim.org 11y ago

Welcome to Eclim – eclim (eclipse + vim)

LiveTheDream
1pts0
phoenix.apache.org 11y ago

Overview – Apache Phoenix

LiveTheDream
4pts0
github.com 11y ago

Paulp/policy – Paul P's fork of scala

LiveTheDream
3pts0
typelevel.org 11y ago

Typelevel Scala and the future of the Scala ecosystem

LiveTheDream
135pts76
stephenramsay.us 11y ago

Life on the Command Line

LiveTheDream
2pts0
brooker.co.za 11y ago

The power of two random choices (load balancing)

LiveTheDream
1pts0
blogs.atlassian.com 11y ago

What's new in Git 2.1

LiveTheDream
2pts0
www.cs.utexas.edu 11y ago

On the foolishness of “natural language programming”

LiveTheDream
132pts152

I’m on a slightly modified version of 0.52.1 which is getting a bit dated but it works well for me even with not officially supported source, like svelte.

In case this thread helps someone else, some errors with —show-repo-map can be solved by setting environment variable PYTHONIOENCODING=utf-8

For me, this has been perplexity.ai. Give it a query, it expands that into multiple queries (possibly chained depending on results) and synthesizes the results into an answer with citations.

I wondered how this could reliably distinguish between a scene cut and a cut to commercial without content hashes and/or program schedules being shared through the network, then realized that 20 years ago was already 2003 and of course home internet was common by then.

Apparently one offline technique was checking for black frames inserted by local stations.

Some more information here: https://en.m.wikipedia.org/wiki/ReplayTV

Update: I got an automated email from Goodreads to download this export. It contains a bunch of json files containing user/usage-related data for my account like request logs, newsfeed updates, kindle logins, site settings, etdc.

You can request your data here https://www.goodreads.com/user/edit?ref=nav_profile_settings

They will send a confirmation email. After you click the link to confirm, it says "We will provide your information to you as soon as we can. Usually, this should take no more than a month."

EDIT: you can export your book database as CSV (shelves, ratings, reviews, etc) here https://www.goodreads.com/review/import (there's a link towards the top to export...disregard that the URL says "import".

FTA: Although the technology is impressive, we’re still a long way from it being deployed in an actual warehouse, especially around humans. That would involve a level of complexity that robots haven't yet mastered.

I am a huge fan of using `make` for this sort of ad hoc data pipeline. The workflow is very natural, as you can play around with each step on the command line and then drop it into the makefile once you get it right..better for reproducibility than search up through terminal history to replay individual lines!

In your example, I would drop the shell and python scripts and simply run:

    watch -n 2 -d "touch a && make"

Tapad | Senior Software Engineers, Data Scientists, Senior Infrastructure Engineers | NYC | Full-time | http://www.tapad.com

We build a probabilistic graph of internet-connected devices based on billions of signals. Lots of Scala-based big data problems to hack on! The realtime systems handle many 100s of thousands of QPS with mere millisecond latency.

How can you detect and remove anomalous data (from botnets perhaps) from datasets that are multiple petabytes in size? Can you write a connected components algorithm that works efficiently at such a scale? Can you write a model that detects individual behaviors within a cluster of device activity?

We run our own datacenters globally, are migrating systems over to Mesos (ooh, shiny), and infrastructure is a first-class critical project, not an afterthought.

All of this happens with a fairly small team of just over 30 engineers, all working together in NYC.

Come join us.

Email toby@tapad.com and say hi.

Tapad, New York, NY

The company that pioneered true cross-device advertising and analytics. Google for "Nielsen Tapad" to see validation :)

Looking for Scala developers with solid CS fundamentals, strong capacity for critical thinking, and experience working with low latency/high throughput systems or big data analytics. You can learn Scala on the job, but a background in Scala or functional programming in general is a major bonus! We've been coding in Scala from the ground up.

Talented co-workers, challenging problems to solve, great office and location are a few things you'll find here. We value technology and productivity and responsibility. We push ourselves to develop quality software using the latest proven tools. You may have seen us at the ny-scala meetup (hosting tonight! see you there), NE Scala Symposium, or Scala Days. Each engineer has a healthy yearly conference budget to facilitate learning.

Contact email is in my profile, or for fun try POSTing your email and a prime number to https://hello.tapad.com/hello (using parameter names "email" and "prime").