Hacker Newsnew | past | comments | ask | show | jobs | submit | jjcm's commentslogin

I'm seeing around 7.9GB of ram, 120 tokens/s on a 6000 pro blackwell.

Regardless of the quality here, I see this as a failed business model. Subscribable libraries of skills / premade assets was effectively what Tailwind was doing, and they weren't able to make it work any longer.

I could see it maybe as a one-time purchase, but $90/yr/user is somthing I'd never grab.


Not to mention my $299 total Tailwind Plus lifetime license (was TailwindUI, then paid the difference for the Plus bundle) has like 100x more stuff in it than $90/yr or $149 lifetime for—check notes—11 UI transitions...

Agree, also why does everything have to be a subscription nowadays baffles me.

Excuse me, please show a little more respect for my profit margins.

> a target release date of 2030

Don't get me wrong, I'm excited for a StarCraft shooter despite it not being StarCraft at all (I enjoy the universe). I wanted StarCraft: Ghost back in the day, and this feels like a taste of that dream.

That said, targeting a release of 2030 is wild to me. Look at the Astra game dev hype that's hitting twitter right now. Yes it's all filtered to the best possible examples / yes people are not one shotting these things, but still it's very clear that LLMs are beginning to be very capable at game dev in a way that doesn't look/feel like ass.

LLMs are going to become more capable over time. I think most of us can agree that eventually they'll be competitive with AAA developers given enough GPU cycles. I would argue that given rate of improvement, we're looking at that point coming in the next year or two.

To emphasize this point, a year ago GPT-5 got released. This was the type of game it would make: https://youtu.be/yTHo7tMborY?t=324

This is the type of game Astra is making: https://youtu.be/GuO_Eo34C8E?t=348

It's just starting to be capable of making assets in blender / unreal. Given this rate of advancement, what is the state going to be like 4 years from now? It really feels a bit like the "travelling to distant stars" problem, where at some point it's faster to wait for better engines than to leave now. I worry that by the time this gets released, the market is going to be flooded with fully custom AAA games that are hyperniche. 4 years out is a long, long time to wait, and it's a longer time to bet on success / market dynamics.


I just read the book Play Nice: The Rise, Fall, and Future of Blizzard Entertainment by Jason Schreier (without having played any of their game or being aware that BlizzCon is around the corner). It seems they'll repeat their own mistakes.

I am getting more and more pessimistic about these studios and their AAA works. Every game is now a gamble with multi year effort with hundreds if not thousands of staff behind it. If a game doesn't sell, and that happens more than once, layoffs are coming. And for Blizzard specifically, it'd better be an online game that's going to be a multi billion dollar cash cow they can keep milking for years. If not, they would not work on the game in the first place.

Of course they have released many great games, but as the history has proven, you can't just keep coming up with new games that are original, fun if not addictive and also make money, especially with the fierce competition today.

Which makes me think I should spend more money/time on indie games than the action/adventure game from big studios that are increasingly repetitive and boring. We need more, smaller studios to give us something new.


Matches my experience. Indie games can be delightful, full of Easter eggs and heart that the dev(s) build for the fun of it. I just played through Hollow Knight for the first time and man, what a great experience.

I'm not a gamedev so I might be wrong, but my friend showed me in 2020 a quick 30 minute tutorial in unity. In this time he made a basic 3d shooter game. The first video is the type of game a teenager would have made 20 years ago and the second one is of course a lot more impressive, but I believe the game engine is doing a lot of the heavy lifting there.

I just started playing with building a game in Godot with Fable 5.1 last weekend and the pace is crazy. Obviously there's a world of difference between what looks playable on my machine and what a AAA studio with the best human artists and developers can put out, and I appreciate that they still exist and put in the first class effort, but... yeah. A 3-year timeline while this is going on seems out of touch with what's going on in the world.

Apparently 6 years is not an uncommon development cycle for AAA right now. For a random example, Ghost of Yotei which is also an open world game, was in development for 4 years. That game is a sequel too, and I think Sucker Punch is probably less dysfunctional than Blizzard at the moment. Still it might have been premature to announce it at this stage.

https://makefaster.dev

I had a bunch of extra Fable credits, so I spent about $10k running autoresearch loops on 200 of the top github repos with frontends to try and speed up the frontend performance. I distilled them down into a leaderboard of the most common wins, and built out an autoresearch loop that takes those learnings and applies them to your repo.

Works with your existing claude/cursor/codex sub in a cute custom TUI.

I just submitted a Show HN for it: https://news.ycombinator.com/item?id=49687032


That's such a great idea. Thanks for doing it!

Really happy for Adam and the rest of the crew - they made a great product for a prior era, and hopefully this is a soft landing for them given their revenue stream dried up.

Tailwind at this point is a coding language and a brand with no direct product. It's a great acquisition when you want mind share and developer love, and are OK with there being no revenue involved. Shopify is a solid match for that.


I use gpt image 2 very heavily for my current project (ai UI design tool). The biggest improvement I'm seeing with this is in speed. I've generated around 50k images with gpt-image-2 via api, the the average latency has held at around 104s.

It's wild how much of a difference this is - images are coming in at around 35-40s. Very noticable, and makes a difference when you're iterating quickly: https://jjcm.org/gpt-image-2.5-speed.mp4


Some UI tests with it:

Warcraft 3 style agentic dev interface: https://image.non.io/cd9ea5cd-8ed7-44e0-ad3f-480ff0e51875.we...

Overall it used the reference images I gave it a bit better than gpt-image-2. I noticed 2 had issues getting the blue button just right. 2.5 nailed it.

A "John Politics" meme site: https://image.non.io/d2922164-fa96-4d07-a141-2febadb02939.we...

Did very well modifying the pose while keeping the appearance of Glenn Powell. gpt-image-2 had a lot of the "fried" look for some of his skin in prior designs I did for johnpolitics.com

A cyberpunk inspired ramen website: https://image.non.io/8d5d8f10-0f0f-4d91-b5ea-33af7538b150.we...

Dark mode sites surfaced the fried look quite a bit in prior models, but this definitely looks better on that front. One thing that looks perhaps worse though is the microglyphs - note the teal lines to the bottom right of the ramen, they're kinda blurry / not straight.

Overall fixed some of the main issues / gripes I had with gpt-image-2


The WC3 one is fun. Definitely feels close to what I remember.

>Warcraft 3 style agentic dev interface

I can forgive the randumb placement of chains but not that sovlless icon of a person from the wow era. wc3 era UI would've used a character portrait.


You have to know that the first one is fine and the next two are pretty horrible, right?

That seems very subjective, can you elaborate? I think the results can be worked with

The 'cyberpunk' one is incredibly bland and uninspired. That may be the goal of the political site, as a form of satire.

But reducing a style with a wealth of cultural material to draw on (literature, fashion, cinema, gaming, industrial design) to 'dark with turquoise/pink highlights, cut corners on rectangles, some random diagonal lines, some token Japanese characters and a lazy neon-style logo' feels like a miss.

If this was from a human designer I was working with, I'd be having a serious conversation about effort.


Love the John Politics page, utterly slick completely bland and phoney. Assume it’s a meme i missed, love it.

I'm confused. Isn't diffui using its own model?

It's both. I have several models loaded into diffui. Which one each prompt uses is determined based on user preference over time. Whichever image is currently at the top of any image node stack marks a win for the model that generated it, I assign each model an ELO score based on that, and I bias the chance each model is selected based on their ELO score. Right now gpt-image-2 is better than my own model, and it services around ~96% of the requests in diffui.

I'll also be adding in microsoft's mai-image-2.6 soon, but I need to update my SOC2 to add MS as a provider before I turn that on for other users. The full list of models in rotation is here: https://image.non.io/d53a9760-8b74-4386-b032-d59da2cd5319.we...


Oh, you're the dev of diffui? It's completely off-topic, but I've notice that RMB -> Download image will download an .png but it's actually a .webp and it can cause issues (e.g. file explorer doesn't show thumbnail correctly).

Indeed I am! Also nice catch - should have a fix up in the next hour. Hit me up with any other reqs - j@diffui.ai

Edit: confirmed that the download should be fixed now - should be properly a png now.


IMO one of the biggest losses of the OpenAI/Cursor breakup will be the loss of OAI models on CursorBench [1]. Their bench has always been one that most-fit my mental model of how good each of these models are. I find AA’s Intelligence index to often be out of alignment with my own subjective evals.

[1] https://cursor.com/evals


It's ability to handle non-90 degree cutouts and shapes for web dev is one of the best I've seen. The vision model on this is VERY capable.

Here's an image design source of truth: https://image.non.io/78f4cd8b-2560-4643-9a51-96a89171f994.we...

And here's the page it build from it: https://image.non.io/e7d3a9e5-f9df-4fd8-b79f-1f90280f978f.we...

Note the flowing svg lines, and how accurately it recreated them. Here's Opus 5 for comparison - you can really see how while Astra really recreated the flow that was in the original design, opus only got the general vibe: https://image.non.io/dfe13de0-4487-431f-8b69-544ff3030dac.we...

One thing I will say is you are paying for quality. That site build cost $24 - extremely non-trivial for a simple frontend.


> That site build cost $24 - extremely non-trivial for a simple frontend.

I would say that $24 is trivial IF that's the final design. The truth is that the cost doesn't leave much room for error or experimentation.


A bespoke design like that a year ago would have cost $200-2000+ between design and development. Hell, it probably still does unless the person wanting the design is already a developer who knows how to prompt.

Everything costs more now than it did a year ago... except for THIS, and we are still complaining that a 90-99% reduction in cost is STILL too expensive. And a 50-75% reduction in time is STILL too long.

We used to have to wait for weeks for a design like that when I worked at a consultancy, and that is a week of salary. For the design, then it got handed off to a front end developer to slice it and get built so the back end developer can hook it up to a CRM. We are talking a month turn around with design, revisions, development, testing, and bug fixing.

It can now be done in a couple of hours for less than a single hour's cost. If it were 10x slower and 10x more expensive, it would STILL be "good deal".


Wdym no room for error or experimation? Changes are even cheaper, and in my experience they are faster and way cheaper with models than with designers.

Plus you can offload a lot of work to a cheaper model

> truth is that the cost doesn't leave much room for error or experimentation

Compared to what?


A cheap Chinese model, of course.

>> The truth is that the cost doesn't leave much room for error or experimentation

Yeah if you are solo developer without budget.

For any business this is nothing, the ROI is massive.


I don't get it. For me it seems Opus was more accurate in terms of for example this small building in the right down corner

I would say opus was in some ways more accurate, but missed the higher level curvature feel of the site that astra picked up on.

It sounds completely trivial and likely I'm wrong here, but could it be that opus saw the reference image squished? That might explain the sharper horizontal curvature


Opus 5 is quite good, and has been the SOTA for this test (by my own subjective comparison) for the last few months. One thing it's always struggled with though is recreating smooth SVG curves/cutouts.

Another way to think about it is you could prompt Astra to put the building back in. You couldn't prompt Opus 5 to get the correct curvature of the line / cutout. That part has always been a huge struggle for models.


They look pretty similar to me, I'm not sure what you're seeing that is so much better.

Also the curves and background lines that separate the right image from the left content

Check the source oft the laser beams.

Ontological discussions aside, I’ll share out my recipe for one of my favorite variations of a Caesar that embraces its origins. It’s one my friends usually request I make/bring.

It’s composed of four main parts, made in this order: the the meat, the croutons, the dressing, and the leaf.

For the meat we’re making a carne asada. You’ll want a flank or a skirt steak, or anything thinner with good marbling. For the marinade use olive oil as the base. I generally do around 2 cups of olive oil for a larger serving. Don’t sweat the exact amounts for a marinade. Add around 10% Apple cider vinegar, 10% red wine vinegar, and 10% lime juice. Use garlic, cumin, black pepper, salt, smoked chipotle peppers (you can also find them in an adobo sauce which works great - skip the sugar if you use the adobo), two spoonfuls of sugar.

Blend all of it together heavily, then pour it over the steak. I like a smoothie blender for this as they tend to work better for these small batches. IMPORTANT: don’t wash the blender, we’ll use the remains each time to influence the subsequent thing. Marinate the steak for at least 2 hours.

Next we’ll make the croutons. Store bought and homemade croutons are very different. For the marinade add a full clove of garlic into our used blender, black pepper, salt, two spoonfuls of nutritional yeast, a thumb sized portion of Parmesan, and rosemary. Add oil so it fills in the gaps in the blender. The goal is a paste, not a liquid. Blend it up. Cut up the bread into slices, then tear it into chunks. The tearing creates better texture. Put the paste in a bowl, smear it all around, dump the bread into, and toss it so it’s coating all the bread. Put it in the oven on a low broil (grill for those in Australia). Careful as it can burn easily. Check it every minute til browns, toss it, then put it back in. Repeat til the outer shell is dry but the inner part is still squishy. They shouldn’t be wet. Remove and chill for 10min before adding them.

For the dressing, half olive oil, 15% lime juice, 15% white wine vinegar, 10% pepitas, 10% Parmesan. This is your base. Add an egg yolk, black pepper, a squirt of mustard. Blend it up, taste it, adjust as needed. I skip the anchovy for the Mexican vibe this salad has as I don’t think it resonates well, but by all means go for it.

Grill the steak high heat.

For the leaf any salad leaf will do, but add radish slices as well.

Toss the leaf and the dressing, top with crouts, steak, pepitas, and Parmesan. Enjoy. Should take you 60-90min of prep.


I've actually had a lot of success with something similar for my startup. I'm a first time founder, and there are so many unknowns when you're getting started. I've been using a Hermes agent to act as my "boss", giving me three tasks to do each day / creating a memory of everything with the business.

It's really, really helpful. So many things I would have completely missed if it weren't for this. I don't think you need this synthetic org setup to get exec help at a small level. It really helped me with fundraising, incorporation, compliance, and just keeping me on track.


> giving me three tasks to do each day

Based on what? How do you judge if those are the right things to work on? As a founder, your number 1 priority is essentially to manage that the company works towards the right thing, but it sounds like to me you're outsourcing that part to a LLM? How do you know what the "right track" is when the track is just what the LLM outlined for you?


> So many things I would have completely missed if it weren't for this.

This is the "unknown unknowns" problem. It sounds like the agent is a reasonably good source of obvious-in-hindsight problems that a experienced founder would have experienced, but a new founder wouldn't easily guess at.

> judge if these are the right things

This problem exists no matter where you get the things from.


That's awesome! In a similar boat, starting something new. Would you mind sharing what customizations you added to Hermes?


I do this with a Gemini instance that is connected to my email and calendar and task list. Every morning I get an email with an agenda. A lot isn't explicitly taken from my task list, but is instead inferred from the contents of emails.

As a mundane example, I had an email from my wife about a kid's swimming party she'd RSVP'd to.

My agenda that morning reminded me to pack, inter alia, sunblock and towels for my new car's seats. So it inferred sunblock from the party mention, and towels from the party mention alongside emails talking about my new car (and some "knowledge" that people take more care to protect new cars vs old cars).


That's more useful than a lot of actual bosses.


This is a fantastic idea and I'd like to add my voice to the choir asking you for more information :) Don't let us rebuild everything from scratch if we don't have to


Would love to see how do you setup the agents, I'm not at the level of director but I think it would be good too for staff engineer to manage the office politics


Executive assistant to intellectual labor seems to be the killer app for LLMs.


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: