Hacker News
Search – A small, fast WebKit browser for macOS
GaryBluto
|next
[-]
jsrozner
|next
|previous
[-]
selectodude
|root
|parent
|next
[-]
packetlost
|root
|parent
|next
[-]
addaon
|root
|parent
[-]
No, testing and code audits also work.
No amount of use of the software will tell you what else the software does beside what is intended (exfiltrate data, etc); testing and code audits both help with that.
packetlost
|root
|parent
[-]
trencedamp
|root
|parent
|next
|previous
[-]
throwaway27448
|root
|parent
|previous
[-]
selectodude
|root
|parent
|next
[-]
drfloyd51
|root
|parent
|previous
[-]
msephton
|root
|parent
|next
|previous
[-]
bs7280
|root
|parent
|next
[-]
When AI can one shot the implementation and testing, it likely only gets 1 round of human testing if you are lucky. Then scale this to an entire AI generated project, the ratio of features to manually run tests is astronomical. In the old days this ratio was inverted, and the tool has been battle tested before reaching any users.
msephton
|root
|parent
[-]
bs7280
|root
|parent
[-]
But - I have found that every single model available is lacking in creating a great UX, even when you are aggressively using playwright in your agent loop. IMO this comes down to a fundamental limitation of how LLMs work. If an app has a button that is too small, or white text on white background, or inconsistent layouts, or non deterministic ux state due to network calls (looking at you spotify).... the agent will not notice or care. A playwright driving agent can click on a button wether its 1px or 1000px wide.
But ultimately, a good ux is about designing for the constraints that we as humans have to live with as mere mortals with all our imperfections. This is not easy to do, and requires a lot of craftmanship by the developers as well as buy in from the suites and middle managers.
In your windows vista example, the issue is almost surely the latter of my previous sentance - middle manager buy in. No one at microsoft leadership gives a shit about the user experience its all kpi seeking nonsense. The user having a good experience is not measured in their spreadsheet.
When you think about it, pretty much every single product we use helps address a problem related to our fundamental human flaws. Im typing this on an expensive ergonomic keyboard because my wrists hurt, youtube premium exists because I dont want to waste my precious time on stuff I dont want to see, we use operating systems with GUIs because our mortal minds cant easily comprehend thousands of lines of text in seconds.
When I use a horrible website I have a very intuitive, almost instinctive reaction similar to how I react to physical pain. An LLM will never experience this, so it will never create a ux as good as a human can on its own.
SrslyJosh
|root
|parent
|next
|previous
[-]
tjpnz
|root
|parent
|next
|previous
[-]
hyperhello
|root
|parent
|previous
[-]
Now endless frameworks and everything make that impossible. So the next step is probably to describe the thing you want to your own AI, in plain English, and have it code it itself.
Of course that will only work until we start using frameworks and everything…sigh.
muhammadusman
|next
|previous
[-]
nexo-v1
|next
|previous
[-]
HelloUsername
|next
|previous
[-]
polycancel
|root
|parent
[-]
frizlab
|root
|parent
[-]
- Resources are 130MiB, including a phishing filter set json file of 18 MiB, phishing hash prefixes of 6 MiB, lots of images and other resources (js, etc.)
- The main binary is 87MiB
- Another binary of 32MiB for “private information removal”
- VPN proxy extension is 17MiB, the VPN binary is 21MiB + 10MiB in another file
- The network protection app extension is 25MiB + another 25 in another file
The rest is frameworks (mostly Lottie and GRDB, in terms of space).
SG-
|next
|previous
[-]
sadly none of that is in this so far.
also boasting about chrome extension support is nice, but at this time people probably want firefox extensions so that UBO can fully run.