HN user

mustaphah

2,156 karma

https://hadid.dev

Posts128
Comments130
View on HN
github.com 6d ago

The Little Book of Reinforcement Learning

mustaphah
213pts26
bear.warrington.ufl.edu 7d ago

Opportunity cost neglect (2009) [pdf]

mustaphah
2pts0
fermatslibrary.com 9d ago

Economic Possibilities for our Grandchildren (1931)

mustaphah
4pts1
twitter.com 18d ago

A field guide to Fable: finding your unknowns

mustaphah
1pts1
www.faros.ai 24d ago

Ten Takeaways from the AI Engineering Report 2026: The Acceleration Whiplash

mustaphah
2pts0
newsletter.kentbeck.com 25d ago

The cost YAGNI was never about

mustaphah
6pts1
www.nber.org 1mo ago

Writing code vs. shipping code [pdf]

mustaphah
3pts0
newsletter.kentbeck.com 1mo ago

Trust Factory

mustaphah
7pts0
www.pnas.org 2mo ago

LLMs pass a standard three-party Turing test

mustaphah
3pts1
hadid.dev 2mo ago

The small sample trap in A/B testing

mustaphah
4pts1
news.ycombinator.com 4mo ago

Tell HN: Claude two rate limits don't know about each other

mustaphah
2pts0
med.stanford.edu 4mo ago

Enhancing gut-brain communication reversed cognitive decline in aging mice

mustaphah
386pts185
metr.org 4mo ago

Many SWE-bench-Passing PRs would not be merged

mustaphah
278pts153
www.tandfonline.com 4mo ago

AGI is an unscientific myth

mustaphah
4pts2
github.com 5mo ago

Web Verbs

mustaphah
1pts0
openai.com 5mo ago

OpenAI's 5-month experiment: building a product with no human-written code

mustaphah
2pts0
arxiv.org 5mo ago

SkillsBench: Benchmarking how well agent skills work across diverse tasks

mustaphah
364pts171
arxiv.org 5mo ago

Evaluating AGENTS.md: are they helpful for coding agents?

mustaphah
232pts161
cursor.com 5mo ago

Curosr: Expanding our long-running agents research preview

mustaphah
3pts0
metr.org 5mo ago

Measuring Time Horizon Using Claude Code and Codex

mustaphah
1pts0
arxiv.org 5mo ago

SWE-ContextBench: context learning benchmark in coding

mustaphah
1pts0
arxiv.org 5mo ago

SWE-AGI: benchmarking spec-driven software construction

mustaphah
1pts1
arxiv.org 5mo ago

Code Formatting Silently Consumes Your LLM Budget

mustaphah
1pts0
agent-trace.dev 5mo ago

Agent Trace by Cursor: open spec for tracking AI-generated code

mustaphah
1pts0
metr.org 5mo ago

METR releases Time Horizon 1.1 with 34% more tasks

mustaphah
1pts0
examine.com 5mo ago

Coffee timing isn't one-size-fits-all

mustaphah
4pts0
blog.kilo.ai 5mo ago

ChatGPT subscription support in Kilo Code

mustaphah
1pts0
www.psypost.org 5mo ago

Imposter Syndrome Predicts Perfectionism

mustaphah
2pts0
www.psypost.org 5mo ago

Motivation acts as a camera lens that shapes how memories form

mustaphah
2pts0
x.com 5mo ago

Claude Code: Merging Slash Commands into Skills

mustaphah
2pts2

Location: Baghdad (UTC+3)

Remote: yes

Willing to relocate: depends

Technologies: TypeScript/Node, Ruby/Rails, Python, Frontend (JS/Dom, React, Tailwind, ...), Microservices & Distributed Systems, REST APIs, GraphQL, RabbitMQ, Pub/Sub, Redis, Postgres/MySQL, Elastic Stack, Prometheus, Splunk, Kubernetes/Docker, Ansible.

Website: https://hadid.dev

Résumé/CV: https://hadid.dev/resume/

GitHub: https://github.com/mhadidg

Email: career+hn @ [my website domain]

---

Hi! Senior software engineer with strong infra/DevOps experience here. I have 8 YOE with a mix of enterprise and startup; 3+ YOE working remotely in a globally distributed team. Looking for a backend or backend-leaning fullstack role with product thinking.

Early in my career, I led the technical side of a workflow automation project at Earthlink - a big local enterprise. Later, I contributed 3+ years to Automattic (US) - the company behind WordPress. I've built and maintained time-sensitive, high-throughput services processing millions of ops daily.

While I'm a technical guy by title, I've worked very closely with business and have decent product development experience.

I do my best on high autonomy, ambiguity, and solving hard problems. I know how to turn vague business needs into systems - I've been doing that for most of my career.

Haidt, in his great book "The Righteous Mind," has been arguing that reasoning evolved not to discover truth but to win arguments. There's a lot of scientific research backing his idea.

Haidt's metaphor is the rider and the elephant: the elephant (intuition) leans, and the rider (reasoning) invents the justification afterward and then defends it like a lawyer, not a truth-seeker.

Intelligence doesn't fix this - it just makes people better at coming up with hard-to-defeat arguments; that explains why smart people disagree all the time.

You can probably catch a big pie of those with simple heuristics to flag suspicious repos for expensive review (human- or AI-based). I did that with public account & repo data, and I believe they can do much more given the amount of private data they have access to.

I'm talking about 10s of repos flagged in a few hours. I don't think the volume would be that big for an expensive review.

Well, my trend detection logic rewards recent stars more than older ones [1]. Recency is an important factor for many custom and public tools that track GitHub trends. I think the bad guys intentionally recreate repos - I actually noticed that.

That being said, they do take action if you report the repo. So I'm guessing good users are doing the heavy lifting here with reporting. I don't believe GitHub is taking enough proactive measures, or maybe they do, but it's not working well, obviously.

https://hadid.dev/posts/github-trends/#growth-based-approach

This is just one flavour of abuse. GitHub does NOT give a shit about the scale of the malware problem.

I've seen so many forms of malware repos working on a GitHub trends newsletter [1], mostly about crypto, NFTs, KMS, and similar stuff.

In the first runs of the project, I was so surprised by tens of malware repos that looked like trending repos. A lot of them share some common traits that made filtering feasible:

- Made by a fresh GitHub user - many created in the past few days.

- The average creation date of Stargazers accounts is very close to the repo creation date. If you take the mean time diff, those bad repos get exposed.

I reported 10s of malware repos, but then I gave up as I felt GitHub was not really doing enough to fight back. I was like... these guys don't seem to care, why should I?

God knows how many people have been abused by these malware repos on GitHub.

---

[1] https://github.com/mhadidg/gh-trends

Is Fable 5 Back? 1 month ago

You want my email voluntarily for the whole purpose of telling me "Hey, Fable is back"?

Everyone would be screaming the moment that happens. No, Thanks!

Aliens.gov 2 months ago

I have a strong feeling the whole thing is distracting the people - the real question is, distracting from what?

Sam Altman and other big figures tend to shape their narratives around their personal and organizational interests. When people were skeptical, they pushed hard into the "God-like AI" narrative. Now that safety concerns are growing and their growth plans are in danger, they're pushing back against what they used to advocate.

Even if they genuinely believe what they’re saying, their perspective is still fundamentally biased and should always be taken with a healthy grain of salt.

I hate to say it, but I'm becoming less and less interested in structured content, and more interested in disorganized, messy content over time. I don't like the thought of how this may end up in a few years for me.

That's a fair clarification.

I never said we have sufficient evidence to act. But "too good to be true" + "singular paper" together can become an unfalsifiable dismissal - by that logic, every important result looks suspicious before it replicates. The interesting question is what priors should update our confidence here.

Stanford/Arc Institute and published in Nature + mechanistic grounding + prior research on gut-brain axis gives me way more confidence than average, but you're right, that's not nearly enough for most, but quite sufficient for me, and surely others with informed priors or a strong motive.

I guess it's because most major disorders and diseases have so many pathways at play that figuring out which one's actually causing the problem at the individual level is just too tricky.

The other thing concerns how potent the effect is to be therapeutic. In many cases, the effect is just marginal to be meaningful.

Yeah, it's a mouse study, but there are tons of human studies backing the whole gut-brain connection. There are even a bunch of books on it [1][2].

What's really cool is that the paper used low-dose capsaicin (just 5 μg/kg injected), and it completely restored hippocampal FOS activity and memory in older mice. Basically, that's the same stuff you get in cayenne pepper supplements - pretty easy to get your hands on.

[1] https://www.goodreads.com/book/show/28837738-the-mind-gut-co...

[2] https://www.goodreads.com/book/show/35210457-the-psychobioti...

[dead] 5 months ago

Well, it's fixed now. Apparently, Claude Code was still up and running!

Even worse, I got contacted through YC Jobs (workatastartup.com) with a message that was basically: "Star, fork, and submit PRs to our open-source repo and we'll review you for a contract."

I immediately realize it's engagement farming + free labor. I said "No thanks."

Got this reply: "(...) I'm looking forward to reviewing your PRs. Feel free to share me any of your questions. (...)"

Apparently, no one read my reply - not even AI. They are automating this shit. It's sad that many fall for it (check their Github repo)

---

Company: Aden (W20)

Contact: Vincent Jiang, Founder

Github: https://github.com/aden-hive/hive