Hacker Newsnew | past | comments | ask | show | jobs | submit | davely's commentslogin

In 2004, I visited New Zealand. It was my first international trip as part of a 6 week long geology field camp. One of the coolest experiences was getting to visit the Fox Glacier and Franz Josef Glacier on the South Island.

Checking this out, I don't see the Fox Glacier or Franz Josef Glacier specifically listed in the search field, but moving the map around, I was able to find the ice fields.

I see that they both will "survive" to 2100, but I wonder in what state. Check out this time lapse video from 2013 to 2018 showing the Franz Josef Glacier's drastic retreat! [3]

[1] https://en.wikipedia.org/wiki/Franz_Josef_Glacier

[2] https://en.wikipedia.org/wiki/Fox_Glacier

[3] https://youtu.be/LUs_20kRHDw?si=k1nJB2z_kn8T--uT&t=17


Yeah, it doesn't seem to make any of the glaciers in NZ, even when they do have names.

Limitation of the data set I guess.


Also one of my fav experiences skydiving over those glaciers. My teeth were so cold!

In my opinion, this is a bit disingenuous.

They were _temporarily_ increased in May by 50% [1]. They continued to extend them through July and August (admittedly, their messaging around this has just been a complete mess and they frequently pushed the deadline back as it approached).

So, now they are giving you a 25% quota increase compared to where things originally stood in May.

So, let me ask you this: assuming you knew that the 50% quota increase was temporary all along, would you then have complained about Anthropic restoring things back to the original limit?

[1] https://www.anthropic.com/news/higher-limits-spacex


Yes, some people will complain about anything (and everything) related to AI. And relentlessly push the most negative interpretation of any datum.


On the contrary, you and Anthropic are being disingenuous by pretending that a usage reduction is actually an increase. Especially when the 20x max plan isn't actually anywhere near 20x, as people have recently realized.


Perhaps! But I do think the vocabulary issue is real and I think LLMs are still sycophantic enough that they won’t really challenge someone or offer alternative ideas on how to implement something unless they explicitly ask.

Interestingly at my work, Claude Code was available before Claude Desktop, so a number of non-technical PMs tried to use it in order to build… anything, with very mixed success.

The “hey guys, check out the website I built with Claude: http://localhost:3000/” joke is real!

In my experience, the whole “the terminal is a scary place” aspect is very real and some non-technical people can feel intimidated by.

I think Claude Code in the desktop app helps alleviate that a bit (perhaps Codex, too, but man what a mess the ‘ol ChatGPT app has become).

But I’m sure there are entire repos of web dev skills that someone could use to put together things with a bit of effort.


> the terminal is a scary place

Isn't the the powerful, unlimited, unopinionated blank LLM text input waiting for your instructions eerily similar to a scary terminal?

WIMP and GUI paradigms are the exact the opposite: intentional dis-empowering, by design restrictions, enumeration of your few possible options. Those feel more constrained therefore safer.


Terminals are scary because it feels insurmountable. What are you supposed to do? If you just type “start python program” it gives you this absurd error that doesn’t make sense. What do you mean start is not in path?

The moment you interact with an LLM it gives you feedback that you’re doing things right. It feels like a gradual climb instead of a series of abrupt jumps. People really don’t like feeling like they don’t know what they’re doing, and the terminal constantly reminds you that you are making mistakes.


Scary but then liberating once you get to know what you have to type in to reach your objective.


Yeah, if you're willing to put in that much effort.

I started on this path literally about as soon as I could read thanks to the family having bought a Commodore 64 for my older siblings, but also perfect timing in that when I got to this age the sibling whose room it was in had just gone off to university.

Most people are not like this, in much the same way that they're not going to read the T&C end-to-end (another thing I've done) or learn enough law to actually understand what those words mean (a step too far even for me).


What an ode to exceptionalism /s

You also think that you are smarter than the people flooding Ceuta streets these days, don't you?


> What an ode to exceptionalism

I've had people criticise me for having had the opportunity to learn in that way, as they did not.

> You also think that you are smarter than the people flooding Ceuta streets these days, don't you?

No, why would I think that? I don't know them, the only thing I can say is in their favour: moving country to better your situation is difficult and them getting as far as they did is a demonstration of putting in a lot of effort of the exact type I praise by default.


> Isn't the the powerful, unlimited, unopinionated blank LLM text input waiting for your instructions eerily similar to a scary terminal?

This is why it has the title (for me currently reading "What can I help with?" but this varies a lot) and the text box itself has the placeholder text "Ask anything". Sometimes I get big friendly suggestions about what to ask it, placed on screen near that text box.

> WIMP and GUI paradigms are the exact the opposite: intentional dis-empowering, by design restrictions, enumeration of your few possible options. Those feel more constrained therefore safer.

I don't think it's constraints, per se: almost nobody looks at the font list and goes "oh no, too many options!"

Rather, GUIs are there to organise your options visually, group them in ways easy to intuitively get. There's a bit of fashion-induced rot here, e.g. I'm old enough to remember when it was always unambiguous when you were looking at a checkbox vs. a radio button, and now there's a blurry middle ground of collections of boxes with ticks in them that act mutually exclusive, but the point of a GUI from a UX POV is not the same as how software in general drifted as it got both more users and more developers and more opinionated managers and middle managers and designers who only cared about shiny rather than usability.


I am old enough to have observed non-tech workers using all kinds of text-based interfaces and it was a real pleasure seeing how old ma's would jiggle numbers on the bc-style TUI in the way that would offset any modern CompSci major.


Me too. But did you see them on their first day using it or their thousandth?

(Also you don't need to be that old. Less than 10 years ago I watched a doctor breeze through some clinical system while I was crawling along constantly referring to the manual)


> WIMP and GUI paradigms are the exact the opposite: intentional dis-empowering

Are you kidding? WIMP and GUI democratized computing!


What "democratized computing" was cheaper computing and VisiCalc, not WIMP and GUI?

But again, VisiCalc is intentionally limited, it's not a all-powerful environment, on purpose. It's all about intentional limitations, making computation easier to reason about.


To put it more bluntly. Computer usage would be low if limited to terminal interactions and there’s no GUI.


Democracy clearly disempowered the monarchs. :)


It did!


With the way subscription models and terms of use have been evolving over time, even if you buy a futuristic robotic assistant outright, you won’t really “own” it.


Excuse me! I hide them in a rats nest under my desk like a civilized person.


Civilized people have a box full of dongles in a cabinet drawer, right next to the box with ethernet cables and the video cable box.


I built something similar awhile back [1] and used OpenAI’s tokenizer playground [2] to recalculate tokens on a giant block of lorem ipsum text. I feel like this gives a much more accurate representation.

[1] https://dave.ly/tools/tokenflow/

[2] https://platform.openai.com/tokenizer


Dear Cliff,

I'm sorry to hear that you must have died right after we visited you in May of 2024 [1]. But I'm glad that you've figured things out and that you are now alive again! I imagine that's a much better state of affairs for you.

Here's to many more years of adventures and quilt making.

Cheers!

[1] https://daveschumaker.net/adventures-in-topology-the-cuckoos...


Writing from this side of death seems to be fairly easy, Dave.

Smiles, -Cliff (who just started designing a new quilt) PS - I quite remember your visit with your daughter a year or so ago)


Haha! Instead, you’ll get a robot that will make you art, music, and tell you stories and you get to toil away cleaning the house.


> With Airdrop you have trivially easy, "just works" sharing with people in proximity.

Funny enough, I encounter so many problems trying to share things via AirDrop with friends, family, and even my own Apple devices that I just tell everyone to install LocalSend and I find that things work better.

I’m not sure why that is, because AirDrop used to work pretty well for me. But it’s been an exercise in frustration more often than not for me.

(Obviously, LocalSend works only as long as everyone is on the same network.)


I've found it very often falls back to sending over your internet connection even if your cell reception sucks. No idea why. People on a previous HN thread talked about solutions


I’ve been on the Claude Code train for a while but decided to try Codex last week after they announced the $100 USD Pro plan.

I’ve been pretty happy with it! One thing I immediately like more than Claude is that Codex seems much more transparent about what it’s thinking and what it wants to do next. I find it much easier to interrupt or jump in the middle if things are going to wrong direction.

Claude Code has been slowly turning into this mysterious black box, wiping out terminal context any time it compacts a conversation (which I think is their hacky way of dealing with terminal flickering issues — which is still happening, 14 months later), going out of the way to hide thought output, and then of course the whole performance issues thing.

Excited to try 4.7 out, but man, Codex (as a harness at least) is a stark contrast to Claude Code.


Do this -- take your coworker's PRs that they've clearly written in Claude Code, and have Codex/GPT 5.4 review them.

Or have Codex review your own Claude Code work.

It then becomes clear just how "sloppy" CC is.

I wouldn't mind having Opus around in my back pocket to yeet out whole net new greenfield features. But I can't trust it to produce well-engineered things to my standards. Not that anybody should trust an LLM to that level, but there's matters of degree here.


I've been using Claude and Codex in tandem ($100 CC, $20 Codex), and have made heavy use of claude-co-commands [0] to make them talk. Outside of the last 1-2 weeks (which we now have confirmation YET AGAIN that Claude shits the fucking bed in the run-up to a new model release), I usually will put Claude on max + /plan to gin up a fever dream to implement. When the plan is presented, I tell it to /co-validate with Codex, which tends to fill in many implementation gaps. Claude then codes the amended plan and commits, then I have a Codex skill that reviews the commit for gaps, missed edge cases, incorrect implementation, missed optimizations, etc, and fix them. This had been working quite well up until the beginning of the month, Claude more or less got CTE, and after a week of that I swapped to $100 Codex, $20 CC plans. Now I'm using co-validation a lot less and just driving primarily via Codex. When Claude works, it provides some good collaborative insights and counter-points, but Codex at the very least is consistently predictable (for text-oriented, data-oriented stuff -- I don't use either for designing or implementing frontend / UI / etc).

As always, YMMV!

[0] https://github.com/SnakeO/claude-co-commands


Some variation of this is the way.

You should not get dependent on one black box. Companies will exploit that dependency.

My version of this is having CC Pro, Cursor Pro, and OpenCode (with $10 to Codex/GLM 5.1) --> total $50. My work doesn't stop if one of these is having overloaded servers, etc. And it's definitely useful to have them cross-checking each other's plans and work.


This more or less mimics a flow that I had fairly good results from -- but I'm unwilling to pay for both right now unless I had a client or employer willing to foot the bill.

Claude Code as "author" and a $20 Codex as reviewer/planner/tester has worked for me to squeeze better value out of the CC plan. But with the new $100 codex plan, and with the way Anthropic seemed to nerf their own $100 plan, I'm not doing this anymore.


> It then becomes clear just how "sloppy" CC is.

Have you done the reverse? In my experience models will always find something to criticize in another model's work.


I have, and in fact models will find things to criticize in their own work, too, so it's good to iterate.

But I've had the best results with GPT 5.4


It cuts both ways. What I usually do these days is to let codex write code, then use claude code /simplify, have both codex and claude code review the PR, then finally manually review and fixup things myself. It's still ~2x faster than doing everything by myself.


I often work this way too, but I'll say this:

This flow is exhausting. A day of working this way leaves me much more drained than traditional old school coding.


100%. On days when I'm sleep deprived (once or twice a week), I fallback to this flow. On regular days, I tend to write more code the old school way and use things things for review.


> One thing I immediately like more than Claude is that Codex seems much more transparent about what it’s thinking and what it wants to do next. I find it much easier to interrupt or jump in the middle if things are going to wrong direction.

I've finally started experimenting recently with Claude's --dangerously-skip-permissions and Codex's --dangerously-bypass-approvals-and-sandbox through external sandboxing tools. (For now just nono¹, which I really like so far, and soon via containerization or virtual machines.)

When I am using Claude or Codex without external sandboxing tools and just using the TUI, I spend a lot of time approving individual commands. When I was working that way, I found Codex's tendency to stop and ask me whether/how it should proceed extremely annoying. I found myself shouting at my monitor, "Yes, duh, go do the thing!".

But when I run these tools without having them ask me for permission for individual commands or edits, I sometimes find Claude has run away from me a little and made the wrong changes or tried to debug something in a bone-headed way that I would have redirected with an interruption if it has stopped to ask me for permissions. I think maybe Codex's tendency to stop and check in may be more valuable if you're relying on sandboxing (external or built-in) so that you can avoid individual permissions prompts.

--

1: https://nono.sh/


There is a new flag for terminal flickering issues:

> Claude Code v2.1.89: "Added CLAUDE_CODE_NO_FLICKER=1 environment variable to opt into flicker-free alt-screen rendering with virtualized scrollback"


Such an interesting choice for a flag name. NO_BUG_PLEASE=1


there is an official codex plugin for claude. I just have them do adversarial reviews/implementations. etc with each other. adds a bit of time to the workflow but once you have the permissions sorted it'll just engage codex when necessary


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: