This is a good question -- I've wanted transparency from image models for a while. One work around is to ask for a "green screen" and to key out the background but it doesn't always work very cleanly.
HN user
niwrad
I've also added a text-to-image versions of the same test here: https://www.cinemodels.ai/benchmark?test=wink&type=text-to-v...
Prompt: "A white female in their 30s winks at the camera with her right eye.
She is standing in a quiet dark green forest, backlit by the sunset during golden hour."
That’s really cool — it makes me happy to hear from someone that appreciates the novelty and feat of his experiment just through perceptual matching! The precision he achieved with such limited tools is honestly mind-blowing.
What are you developing a spectral CMS for? Is it for lighting or materials or something imaging related?
Absolutely — these old papers are fascinating. I was also surprised by how much insight is packed into them, especially considering how non-trivial some of it is to unpack. My professor and I spent quite a bit of time reconstructing how the experiment worked from the limited figures.
Thank you!
Maxwell's "color wheel" experiment (https://cudl.lib.cam.ac.uk/view/PH-CAVENDISH-P-02000/1) is more commonly known, which preceded this experiment I wrote about. (Which by the way is also a very clever experiment which blends color by spinning a wheel with different ratios of primary colors).
Here's a link to Maxwell's original paper: https://royalsocietypublishing.org/doi/10.1098/rstl.1860.000...
You can see some diagrams of his original apparatus that I worked off of in the last two pages!
An audience-driven GenAI rom-com w/ Daily Episodes.
How We Met – https://how-we-met.c47.studio/
Each day, I create a new 30-second episode based on the plot direction voted on by the audience the day before.
I'm trying to see how far the latest Video GenAI can go with narrative content, especially episodics. I'm also curious what community-driven narratives look like!
For the past week, I've been tinkering mostly with Runway, Midjourney, and Suno for the video content. My co-creator vibe coded the platform on Lovable.
Thank you for the encouraging words! I’m glad you enjoyed it.
I was genuinely confused when I saw the difficulty in getting these IC cards in July. I’m even more confused to see them completely stop the sales all together.
These cards are ubiquitous in Japan so I’m very confused how this can happen.
Can someone that has some understanding of semiconductor supply chain explain what’s going on here?
The chips inside these cards surely don’t seem like the high-end chips that the AI crowds are going around.
I ended up subscribing to the $60 plan, mainly to get access to the Stealth Mode. I used about ~4h of fast time during the project. With that said, I could have created this with the $10 plan (3.3 hours) if I had to.
Thanks for your kind comments!
I'm glad you pointed out the music. The music was also AI generated with a tool called AIVA [1]. I'd never composed a piece of music before, and I was pretty surprised by what I could "create". I spent 30~60 minutes max creating the score.
Some parts of their product still feel janky, but as an overall concept, it's quite fascinating. One of the interactions I enjoyed was that AIVA creates scores with different tracks (layers). So I was able to edit tracks I don't like (e.g., change a Piano track to Brass) or have AIVA completely regenerate certain sections of the score (e.g., redo the bridge, regenerate the chorus sections).
One difference from Midjourney is that there's no text-based prompting. Instead, you "prompt" through music inspiration.
I feel that Midjourney v5 really lets you explore different worlds.
One recent feature the guide missed is the permutation and repeat features [1]. They're quite helpful for power users that want to explore multiple styles quickly.
Last week I tried putting together a short film using GPT-4 and Midjourney v5. I was stunned by the cinematic frames Midjourney v5 was able to create:
I (human) wrote the prompts for Midjourney, though.