15 ms·
Schedule tasks on the web
- pastel8739 6mo agoIs this free? I don’t see pricing info. I guess just a way to make you forget that you’re spending money on tokens?
- weird-eye-issue 6mo agoYou don't spend money on tokens. It is a subscription.
- arjie 6mo agoWhat's the per-unit-time compute cost (independent of tokens)? Compute deadline etc.? They don't charge for the Cloud Environment https://code.claude.com/docs/en/claude-code-on-the-web#cloud-environment https://code.claude.com/docs/en/claude-code-on-the-web#cloud... currently running?
- nickandbro 6mo agoI feel like we are just inching closer and closer to a world where rapid iteration of software will be by default. Like for example a trusted user makes feedback -> feedback gets curated into a ticket by an AI agent, then turned into a PR by an Agent, then reviewed by an Agent, before being deployed by an Agent. We are maybe one or two steps from the flywheel being completed. Or maybe we are already there.
- theredbeard 6mo agoWe haven’t been inching closer to users writing a half-decent ticket in decades though.
- edf13 6mo agoOr perhaps we end up where all software is self evolving via agents… adjusting dynamically to meet the users needs.
- PeterStuer 6mo agoThe "user" being the one that's in charge of the AI, not the person on the receiving end.
- charcircuit 6mo agoThen sets up telemetry and experiments with the change. Then if data looks good an agent ramps it up to more users or removes it.
- jvuygbbkuurx 6mo agoTusted user like Jia Tan.
- tossandthrow 6mo agoI think the Ai agent will directly make a PR - tickets are for humans with limited mental capacity. At least in my company we are close to that flywheel.
- Gigachad 6mo agoThe agents have even more limited capacity
- _puk 6mo agoTickets need to exist purely from a governance perspective. Tickets may well not look like they do now, but some semblance of them will exist. I'm sure someone is building that right now. No. It's not Jira.
- tossandthrow 6mo agoYes, so my point is that PRs act as that governance layer - with preview environments, you can see the complexity and risk of the change etc.
- yieldcrv 6mo agoWe do feedback to ticket automatically We dont have product managers or technical ticket writers of any sort But us devs are still choosing how to tackle the ticket, we def don't have to as I’m solving the tickets with AI. I could automate my job away if I wanted, but I wouldn't trust the result as I give a degree of input and steering, and there’s bigger picture considerations its not good at juggling, for now
- MattGaiser 6mo agoI am already there with a project/startup with a friend. He writes up an issue in GitHub and there is a job that automatically triggers Claude to take a crack at it and throw up a PR. He can see the change in an ephemeral environment. He hasn't merged one yet, but it will get there one day for smaller items. I am already at the point where because it is just the two of us, the limiting factor is his own needs, not my ability to ship features.
- jondwillis 6mo agoWhy doesn’t he merge them?
- MattGaiser 6mo agoHe is not technical but a product guy, so he still wants me to check it over.
- m00x 6mo agoMust be nice working on simple stuff.
- chatmasta 6mo agoI love everything about this direction except for the insane inference costs. I don’t mind the training costs, since models are commoditized as soon as they’re released. Although I do worry that if inference costs drop, the companies training the models will have no incentive to publish their weights because inference revenue is where they recuperate the training cost. Either way… we badly need more innovation in inference price per performance, on both the software and hardware side. It would be great if software innovation unlocked inference on commodity hardware. That’s unlikely to happen, but today’s bleeding edge hardware is tomorrow’s commodity hardware so maybe it will happen in some sense. If Taalas can pull off burning models into hardware with a two month lead time, that will be huge progress, but still wasteful because then we’ve just shifted the problem to a hardware bottleneck. I expect we’ll see something akin to gameboy cartridges that are cheap to produce and can plug into base models to augment specialization. But I also wonder if anyone is pursuing some more insanely radical ideas, like reverting back to analog computing and leveraging voltage differentials in clever ways. It’s too big brain for me, but intuitively it feels like wasting entropy to reduce a voltage spike to 0 or 1.
- eksu 6mo agoThis is the wrong way to see it. If a technology gets cheaper, people will use more and more and more of it. If inference costs drop, you can throw way more reasoning tokens and a combination of many many agents to increase accuracy or creativity and such.
- gf000 6mo ago> throw way more reasoning tokens and a combination of many many agents to increase accuracy or creativity and such. But this is just not true, otherwise companies that can already afford such high prices would have already outpaced their competitors.
- spacebanana7 6mo agoNo company at the moment has enough money operate with 10x the reasoning tokens of their competitors because they're bottlenecked by GPU capacity (or other physical constraints). Maybe in lab experiments but not for generally available products. And I sense you would have to throw orders of magnitude more tokens to get meaningfully better results (If anyone has access to experiments with GPT 5 class models geared up to use marginally more tokens with good results please call me out though).
- bredren 6mo agoWhat you're describing is absolutely where we're headed. But the entire SWE apparatus can be handled. Automated A/B testing of the feature. Progressive exposure deployment of changes, you name it.
- slopinthebag 6mo agoWhat kind of software are people building where AI can just one shot tickets? Opus 4.6 and GPT 5.4 regularly fail when dealing with complicated issues for me.
- thin_carapace 6mo agoi dont see anyone sane trusting ai to this degree any time soon, outside of web dev. the chances of this strategy failing are still well above acceptable margins for most software, and in safety critical instances it will be decades before standards allow for such adoption. anyway we are paying pennies on the dollar for compute at the moment - as soon as the gravy train stops rolling, all this intelligence will be out of access for most humans. unless some more efficient generalizable architecture is identified.
- slopinthebag 6mo agoEven in webdev it rots your codebase unchecked. Although it's incredibly useful for generating UI components, which makes me a very happy webslopper indeed.
- thin_carapace 6mo agoim grateful to have never bothered learning web dev properly, it was enlightening witnessing chat gpt transform my ten second ms paint job into a functional user interface
- m00x 6mo agoSeveral fintechs like Block and Stripe are boasting thousands of AI-generated PRs with little to no human reviews. Of course it's in the areas where it doesn't matter as much, like experiments, internal tooling, etc, but the CTOs will get greedy.
- slopinthebag 6mo agoI don't think anybody is doubting its ability to generate thousands of PR's though. And yes, it's usually in the stuff that should have been automated already regardless of AI or not.
- eranation 6mo agoUm, we are already there...
- eranation 6mo agoNot sure why the downvote, I'm seeing this happening...
- Leptonmaniac 6mo agoI think that as a user I'm so far removed from the actual (human) creation of software that if I think about it, I don't really care either way. Take for example this article on Hacker News: I am reading it in a custom app someone programmed, which pulls articles hosted on Hacker News which themselves are on some server somewhere and everything gets transported across wires according to a specification. For me, this isn't some impressionist painting or heartbreaking poem - the entity that created those things is so far removed from me that it might be artificial already. And that's coming from a kid of the 90s with some knowledge in cyber security, so potentially I could look up the documentation and maybe even the source code for the things I mentioned; if I were interested.
- slopinthebag 6mo agoArt is and has always been about the creator.
- vntok 6mo agoTake a walk in any museum, I'm pretty sure you'll react to some of the art displayed there and find it cool before you read the name of the artist.
- rjknight 6mo agoIt's not that you know the artist first and then say "this art is cool because I like the artist". The art is the means by which you know the artist. The more of their works you encounter, the closer you get to understanding the artist and what they are trying to communicate.
- egeozcan 6mo agoDive into a forest, you'll find a couple of cool trees. Art isn't about being cool. Art is about context. When I tell people that art cannot be unpolitical, they react strongly, because they think about the left/right divide and how divided people are, where art is supposed to be unifying. But art is like movement, you need an origin and a destination. Without that context, it will be just another... thing. Context makes it something.
- eru 6mo agoInstead of having a trusted user, you can also do statistics on many users. (That's basically what A/B testing is about.)
- tuo-lei 6mo agoThe missing piece for me is post-hoc review. A PR tells me what changed, but not how an AI coding session got there: which prompts changed direction, which files churned repeatedly, where context started bloating, what tools were used, and where the human intervened. I ended up building a local replay/inspection tool for Claude Code / Cursor sessions mostly because I wanted something more reviewable than screenshots or raw logs.
- hyperionultra 6mo ago"Trusted user" also can be an Agent.
- heavyset_go 6mo agoFeedback loops like that would be an exercise in raising garbage-in->garbage-out to exponential terms. It's the "robots will just build/repair themselves" trope but the robots are agents
- TeMPOraL 6mo agoYes. Next they'll want nanobots that build/repair themselves. Oh wait. That's already here and is working fine.
- heavyset_go 6mo agoTurns out we were the nanobots all along
- mindwok 6mo agoI think Anthropic will launch backend hosting off the back of their Bun acquisition very soon. It makes sense to basically run your entire business out of Claude, and share bespoke apps built by Claude code for whatever your software needs are.
- pxtail 6mo ago100% its going to happen - also OpenAI will do same, there were already rumors about them building internal "github" which is stepping stone for that Also it is requirement for completing lock-in - the dream for these companies.
- overfeed 6mo ago> I feel like we are just inching closer and closer to a world where rapid iteration of software will be by default. There's a lots of experimentation right now, but one thing that's guaranteed is that the data gatekeepers will slam the door shut[1] - or install a toll-booth when there's less money sloshing about, and the winners and losers are clear. At some point in the future, Atlassian and Github may not grant Anthropic access to your tickets unless you're on the relevant tier with the appropriate "NIH AI" surcharge. 1. AI does not suspend or supplant good old capitalism and the cult of profit maximization.
- shafyy 6mo agoHaha sure, let's just let every user add their feedback to the software.
- andy_ppp 6mo agoUsers are often incorrect about what the software should actually be doing and don’t see the bigger picture.
- backscratches 6mo agoIn the past three weeks a couple of projects I follow have implemented AI tools with their own github accounts which have been doing exactly this. And they appear to be doing good work! Dozens of open issues iterated, tested and closed. At one point i had almost 50 notification for one projects backlog being eradicated in 24 hours. The maintainer reviewed all of it and some were not merged.
- jwpapi 6mo agoI just don’t see it coming. I was full on that camp 3 months ago, but I just realize every step makes more mistakes. It leads into a deadlock and when no human has the mental model anymore. Don’t you guys have hard business problems where AI just cant solve it or just very slowly and it’s presenting you 17 ideas till it found the right one. I’m using the most expensive models. I think the nature of AI might block that progress and I think some companies woke up and other will wake up later. The mistake rate is just too high. And every system you implement to reduce that rate has a mistake rate as well and increases complexity and the necessary exploration time. I think a big bulk of people is of where the early adaptors where in December. AI can implement functional functionality on a good maintained codebase. But it can’t write maintable code itself. It actually makes you slower, compared to assisted-writing the code, because assisted you are way more on the loop and you can stop a lot of small issues right away. And you fast iterate everything• I’ve not opened my idea for 1 months and it became hell at a point. I’ve now deleted 30k lines and the amount of issues I’m seeing has been an eye-opening experience. Unscalable performance issues, verbosity, straight up bugs, escape hatches against my verification layers, quindrupled types. Now I could monitor the ai output closer, but then again I’m faster writing it myself. Because it’s one task. Ai-assisted typing isn’t slower than my brain is. Also thinking more about it FAANG pays 300$ per line in production, so what do we really trying to achieve here, speed was never the issue.A great coder writes 10 production lines per day. Accuracy, architecture etc is the issue. You do that by building good solid fundamental blocks that make features additions easier over time and not slower
- EdgeNRoots 6mo ago[dead]
- aspenmartin 6mo agoI think this sounds like a true yet short sighted take. Keep in mind these features are immature but they exist to obtain a flywheel and corner the market. I don’t know why but people seem to consistently miss two points and their implications - performance is continuing to increase incredibly quickly, even if you rightfully don’t trust a particular evaluation. Scaling laws like chinchilla and RL scaling laws (both training and test time) - coding is a verifiable domain The second one is most important. Agent quality is NOT limited by human code in the training set, this code is simply used for efficiency: it gets you to a good starting point for RL. Claiming that things will not reach superhuman performance, INCLUDING all end to end tasks: understanding a vague business objective poorly articulated, architecting a system, building it out, testing it, maintaining it, fixing bugs, adding features, refactoring, etc. is what requires the burden of proof because we literally can predict performance (albeit it has a complicated relationship with benchmarks and real world performance). Yes definitely, error rates are too high so far for this to be totally trusted end to end but the error rates are improving consistently, and this is what explains the METR time horizon benchmark.
- dominotw 6mo agoI dont mean this as a shade but ppl who are not coders now seem to think "coding is now solved" and seem to be pushing absurd ideas like shipping software with slack messages. These ppl are often high up in the chain and have never done serious coding. Stripe is apparently pushing gazzaliion prs now from slack but their feature velocity has not changed. so what gives? how is that number of pr is now the primary metric of productivity and no one cares about what is being shipped or if we are shipping product faster. Its total madness right now. Everyone has lost their collective minds.
- rkomorn 6mo agoI ask myself the same question. I'm not seeing the apps, SaaS, and other tools I use getting better, with either more features or fewer bugs. Whatever is being shipped, as an end user, I'm just not seeing it.
- dominotw 6mo agocto and ceo are now feeling insane pressure to show how they are using ai but its not evident in output. So now they've resorted to blabbering publicly about prs, lines of code ect to save face. And ofcourse ppl giving them voice and platform have their own agendas that prevent them from asking "so what exactly have you shipped stripe from million pr/day". Its baffling to see these comments on hacknernews though. I guess you have to prove that you are not a luddite by making "ai forward" predictions and show that you "get it"
- duped 6mo agoI think a lot of SWE roles are really bullshit jobs (1) and these have been particularly susceptible to getting sniped with AI tools. (1) https://en.wikipedia.org/wiki/Bullshit_Jobs https://en.wikipedia.org/wiki/Bullshit_Jobs
- obastani 6mo agoI don't know if this is the future, but if it is, why bother building one version of the software for everyone? We can have agents build the website for each user exactly the way they want. That would be the most exciting possibility to come out of AI-generated software.
- bwestergard 6mo ago"why bother building one version of the software for everyone?" So one user's experience is relevant to another, so they can learn from one another?
- lancekey 6mo agoHa I just SPECed out a version of this. I have a simple static website that I want a few people to be able to update. So, we will give these 3 or 4 trusted users access to an on-site chat interface to request updates. Next, a dev environment is spun up, agent makes the changes, creates PR and sends branch preview link back to user. Sort of an agent driven CMS for non-technical stakeholders. Let’s see if it works.
- EastLondonCoder 6mo agoI think some type of tickets can be done like this but your trusted user assumption does a lot of work here. Now I don't see this getting better than that with the current architecture of LLMs, you can do all sorts of feedback mechanisms which helps but since LLMs are not conscious drift is unavoidable unless there is a human in the loop that understands and steers what's going on. But I do think even now with certain types of crud apps, things can be largely automated. And that's a fairly large part of our profession.
- fatata123 6mo ago[dead]
- eerikkivistik 6mo agoI know a company already operating like this in the fintech space. I foresee a front page headline about their demise in their future.
- gowthamgts12 6mo agointeresting to see feature launches are coming via official website while usage restrictions are coming in with a team member's twitter account - https://x.com/trq212/status/2037254607001559305 https://x.com/trq212/status/2037254607001559305. also, someone rightly predicted this rugpull coming in when they announced 2x usage - https://x.com/Pranit/status/2033043924294439147 https://x.com/Pranit/status/2033043924294439147
- stingraycharles 6mo agoTo me it makes perfect sense for them to encourage people to do this, rather than eg making things more expensive for everyone. The same as charging a different toll price on the road depending on the time of day.
- tyre 6mo agoIf you read the replies to the second, you’ll see an engineer on Claude Code at Anthropic saying that it is false. Someone spread FUD on the internet, incorrectly, and now others are spreading it without verifying.
- hobofan 6mo agoAnd if you look closely at the usernames, you see that the same engineer from link 2 that said "nah it’s just a bonus 2x, it’s not that deep" (just two week ago) is now saying "we're going to throttle you during peak hours" (as predicted). Yes, it was FUD, but ended up being correct. With the track record that Anthropic has (e.g. months long denial of dumbed down models last year, just to later confirm it as a "bug"), this just continues to erode trust, and such predictions are the result of that.
- browningstreet 6mo agoAnthropic fixing that bug way faster than Apple fixing iOS keyboard "bug". Anthropic even acknowledged it, Apple gave us the silent treatment for years. I'm not sure it's a rug pull when their stats show 7% and 2% subscription-level impacts. We're back in the ISP days, and they never said unlimited.
- iBelieve 6mo agoLooks like I'm limited to only 3 cloud scheduled tasks. And I'm on the Max 20x plan, too :( "Your plan gets 3 daily cloud scheduled sessions. Disable or delete an existing schedule to continue." But otherwise, this looks really cool. I've tried using local scheduled tasks in both Claude Code Desktop and the Codex desktop app, and very quickly got annoyed with permissions prompts, so it'll be nice to be able to run scheduled tasks in the cloud sandbox. Here are the three tasks I'll be trying: Every Monday morning: Run `pnpm audit` and research any security issues to see if they might affect our project. Run `pnpm outdated` and research into any packages with minor or major upgrades available. Also research if packages have been abandoned or haven't been updated in a long time, and see if there are new alternatives that are recommended instead. Put together a brief report highlighting your findings and recommendations. Every weekday morning: Take at Sentry errors, logs, and metrics for the past few days. See if there's any new issues that have popped up, and investigate them. Take a look at logs and metrics, and see if anything seems out of the ordinary, and investigate as appropriate. Put together a report summarizing any findings. Every weekday morning: Please look at the commits on the `develop` branch from the previous day, look carefully at each commit, and see if there are any newly introduced bugs, sloppy code, missed functionality, poor security, missing documentation, etc. If a commit references GitHub issues, look up the issue, and review the issue to see if the commit correctly implements the ticket (fully or partially). Also do a sweep through the codebase, looking for low-hanging fruit that might be good tasks to recommend delegating to an AI agent: obvious bugs, poor or incorrect documentation, TODO comments, messy code, small improvements, etc. I ran all of these as one-off tasks just now, and they put together useful reports; it'll be nice getting these on a daily/weekly basis. Claude Code has a Sentry connector that works in their cloud/web environment. That's cool; it accurately identified an issue I've been working on this week. I might eventually try having these tasks open issues or even automatically address issues and open PRs, but we'll start with just reports for now.
- NuclearPM 6mo ago0 7 * * 1-5 ANTHROPIC_API_KEY=sk-... /path/to/claude-cron.sh /path/to/repo >> ~/claude-reports.md 2>&1 Seems trivial.
- 6mo ago
- lucgagan 6mo agoHere goes my project.
- hydroweaver87 6mo agoWhat were you working on?
- rhubarbtree 6mo agoBetter idea. Watch online feedback on this feature. Then implement things users want. Go niche. Join the forum and help them use Claude to its limits. Then be the next step for power users.
- pxtail 6mo agoWelcome to Amazon playbook replayed again, most useful, profitable and popular use-cases will implemented by platform - and they will do it ruthlessly and quickly as money needs to be recouped.
- codybontecou 6mo agoDon't bet against the models and their providers becoming stagnant. Build with the idea that they will continue to improve.
- zmmmmm 6mo agoi'm missing something basic here .... what does it actually do? It executes a prompt against a git repository. Fine - but then what? Where does the output go? How does it actually persist whatever the outcome of this prompt is? Is this assuming you give it git commit permission and it just does that? Or it acts through MCP tools you enable?
- tossandthrow 6mo agoWe use to do do automated sec audits weekly on the code base and post the result on slack
- zmmmmm 6mo agoso is slack posting an MCP tool it has? or a skill it just knows?
- tossandthrow 6mo agoIn Claude it is a "connector" which is essentially an mcp tool.
- jngiam1 6mo agoMCP tools. We're doing some MCP bundling and giving it here, pretty cool stuff.
- jngiam1 6mo agoThis is powerful. Combined with MCPs, you can pretty much automate a ton of work.
- esperent 6mo agoCan you give some examples?
- adobrawy 6mo agoThat feature was silent launched about week ago for me. I use it to: - perform review of latest changes of code to update my documentation (security policies, user documentation etc.) - perform review to latest changes of code, triage them, deduplicate and improve code - I review them, close them with comments for over-engoneering / add review for auto-fix - perform review of open GitHub issue with label, select the one with highest impact, comment with rationale, implement it and make pull request - I wake up and I have a few pull request to fix issues that I can approve /finish in existing Claude Code thread I want also use it to: - review recent Sentry issues, make GitHub issues for the one with highest priority, make pull request with proposed fix - I can just wake up and see that some crash is ready to be resolved Limit of 3 scheduled jobs is pretty impactful, but playing with it give me a nice idea on how I can reduce my manual work.
- nerptastic 6mo agoMaybe I’m missing the point but we have some of these implemented without the tool - the only one that needs an API key is the log scraping. It’s been surprisingly cheap and if we want to swap models we can.
- adobrawy 6mo agoOf course, this scheduled can be implemented on top of existing tools. However, it's incredibly convenient to have a UI where: - you can easily edit the prompt - you can see the prompt execution history - you don't need any infrastructure to orchestrate it - it works even when your computer is off Once you start using it, it turns out to be very convenient.
- mkagenius 6mo agoThis is a bit restrictive, doesn't take screenshots. So you can't "say take screenshots of my homepage and send it to me via email" It doesnt allow egress curl, apart from few hardcoded domains. I have created Cronbox in the cloud which has a better utility than above. Did a "Show HN: Cronbox – Schedule AI Agents" a few days back. https://cronbox.sh https://cronbox.sh and a pelican riding a bicycle job - https://cronbox.sh/jobs/pelican-rides-a-bicycle?variant=terminal https://cronbox.sh/jobs/pelican-rides-a-bicycle?variant=term...
- chopete3 6mo agoClaude is moving fast. https://grok.com/tasks https://grok.com/tasks Grok has had this feature for some time now. I was wondering why others haven't done it yet. This feature increases user stickiness. They give 10 concurrent tasks free. I have had to extract specific news first thing in the morning across multiple sources.
- MeetRickAI 6mo ago[flagged]
- simianwords 6mo agoI remember when I tried to set something up with the ChatGPT equivalent like "notify me only if there are traffic disruptions in my route every morning at 8am" and it would notify me every morning even if there was no disruption.
- scottmcdot 6mo agoMe too. It doesn't have ability to alert only on true positive. I has to also alert on true negative. So dumb
- worldsayshi 6mo agoThis doesn't seem to hard to solve except for the ever so recurring llm output validation problem. If the true positive is rare you don't know if the earthquake alert system works until there's an earthquake.
- g3f32r 6mo ago... just force the data into a structured format, then use "hard code" on the structure. "Generate the following JSON formatted object array representing the interruptions in my daily traffic. If no results, emit []. Send this at 8am every morning. {some schema}. Then run jsonreporter.py" Then just let jsonreporter.py discriminate however it likes. Keep the LLMs doing what they are good at, and keep hard code doing what it's good at.
- theredbeard 6mo agoThis is because for some reason all agentic systems think that slapping cron on it is enough, but that completely ignores decades of knowledge about prospective memory. Take a look at https://theredbeard.io/blog/the-missing-memory-type/ https://theredbeard.io/blog/the-missing-memory-type/ for a write-up on exactly that.
- alexhans 6mo agoWhy not set your own evals and something like pi-mono for that? https://github.com/badlogic/pi-mono/ https://github.com/badlogic/pi-mono/ You'll define exactly what good looks like.
- PeterStuer 6mo agoIs only Github supported as a repository?
- monkeydust 6mo agoI do feel people will end up using this for things where a deterministic rule could be used - more effective, faster and cheaper. See this starting to happen at work...'We need AI to solve X....no you don't"
- TeMPOraL 6mo agoMaybe. The problem of "execute task on a cron" is something I've noticed the industry seems to refuse to solve in general, as if intentionally denying this capability for regular people. Even without AI, it's the most basic block of automation, and is always mysteriously absent from programs and frameworks (at least at the basic level). AI only makes it more useful on "then" side, but reliable cron on "if" side is already useful.
- monkeydust 6mo agoAgree. How would you solve this in general, what would be the ingredients? People use things like zapier, n8n, node-red to achieve this today but in many cases are overkill.
- bshimmin 6mo agoHonestly, you just need cron (and Ruby/Python/bash/whatever) on an EC2. It's not very fashionable, but it works, will continue to work forever, and costs hardly anything.
- maccard 6mo agoTo use an example in the article, what does > Analyzing CI failures overnight and surfacing summaries Look like on ec2 with python? Because with Claude, it’s that prompt, and with your solution it’s infra + security groups + multiple APIs + whatever code you actually write
- duped 6mo agoI would suggest the prompt is an example of garbage in that's going to produce garbage out. Sitting down to confront the problem you're solving will show this, while Claude is going to happily spit out what looks like a plausibly functional system. So for example the only "analysis" of CI failures are which systems failed and who/what committed the changes to those things. The only way AI would help me here is if the system was so jank that the sole primitive i can use is textual analysis of log files. Which granted is probably real for a lot of software firms, but I really hope I have better build and test infrastructure than that.
- javiercr 6mo agoI've recently switched from GitHub Copilot Pro to Claude Code Max (20x). While Claude is clearly superior in many aspects, one area where it falls short is remote/cloud agents. Yesterday, I spent the entire day trying to set up "Claude on the web" for an Elixir project and eventually had to give up. Their network firewall kept killing Hex/rebar3 dependency resolution, even after I selected "full" network access. The environment setup for "on the web" is just a bash script. And when something goes wrong, you only see the tail of the log. There is currently no way to view the full log for the setup script. It's really a pain to debug. The Copilot equivalent to "Claude on the web" is "GitHub Copilot Coding Agents," which leverages GitHub Actions infrastructure and conventions (YAML files with defined steps). Despite some of the known flaws of GitHub Actions, it felt significantly more robust. "Schedule task on the web" is based on the same infrastructure and conventions as "Claude on the web", so I'm afraid I'm gonna have the same troubles if I want to use this.
- mememememememo 6mo agoThe PHP script from a cron tab is back!
- j1000 6mo agolmao
- maxbeech 6mo ago[dead]
- hirako2000 6mo agoOh my, did Anthropic invent Cron jobs as a service? It's a game changer. Edit: my mistake. It's inferior to a Cron job. If my repos happen to be self hosted with Forgejo or codeberg, then it won't even work. If I concede to use GitHub though I don't have to set up any env variables. Schedules lock-in, all over the web.
- no_shadowban_3 6mo ago[dead]
- TeMPOraL 6mo agoYou jest, but for some reason the industry stubbornly refuses to solve the "cron job as a service" problem for end-users, whether on the web or in the OS. I feel this is rooted in problems that extend beyond computing. Regular people are not allowed to automate things in their life. Consider that for most people, the only devices designed to allow unattended execution off a timer are a washing machine, some ovens and dishwashers, and an alarm clock (also VCRs in the previous era). Anything else requires manual actuation and staying in a synchronous loop.
- hirako2000 6mo agoThere is nothing to solve. It's already there, a VPS, a container platform, just push your script and schedule it. Of course a provider can offer convenient shortcuts, but at the cost of getting tied into their ecosystem. Anthropic is clearly battling an existential threat: what happens when our paying users figure out they can get a better and cheaper model elsewhere.
- TeMPOraL 6mo ago> what happens when our paying users figure out they can get a better and cheaper model elsewhere. They solved that with subscriptions. For end-users (and developers using AI for coding), it makes no sense to go for pay-as-you-go API use, as anything interesting will burn more than the monthly subscription worth of $$$ in API costs in few hours to days.
- throwatdem12311 6mo agoSo this is basically just Anthropic’s version of Open Claw that they manage for you and you pay them.
- commers148 6mo ago[flagged]
- dbvn 6mo agoit would be easier to use claude to write a cronjob that does the same thing for you but accurately
- qznc 6mo agoAnd yet it probably covers 90% of what people use OpenClaw for.
- 0898 6mo agoOne interesting restriction is that it won’t do anything with people’s faces. I run conferences and I like to have photos of delegates on the page so you can see who else is attending. I wanted to automate this by having Claude go to the person’s LinkedIn profile and save the image to the website. But it seems it won’t do that because it’s been instructed not to.
- wslh 6mo agoLinkedIn already employs anti-scraping measures, so I'd expect a lot of users to get flagged. That's not unique to LinkedIn but what is somewhat unique is the strong linkage to real world identities, which raises the cost of Sybil attacks on personal networks with high trust.
- jFriedensreich 6mo agoWe need to fight model providers trying to own memory, workflows and tooling. Don't give them an inch more of your software than needed even if there is a slight inconvenience setting up.
- sharemywin 6mo agoI wish there was a company that was easy to use but wouldn't sell out in this arena.
- solaceb 6mo agohi, I don’t normally promote here, but I feel compelled to ask if you’d like to test my thing. it’s a personal agent / API for creating and managing background cloud agents that I’m 100% committed to keeping open source & accessible as an alternative platform to putting all your eggs in one basket. there is also a desktop app and expanding the api to involve storage. kind of like agentic dropbox that can also do coding and has a full computer and ability to spin up N agents https://tinyfat.com https://tinyfat.com
- kizashi 6mo agoVery much like the idea. Thanks for sharing. Noticed that you are pushing this fully anonymously and wanted to chat with you regarding a project that I’m building. Mind contacting me on the address in my profile?
- Bnjoroge 6mo agothats like looking for a unicorn.
- arrowleaf 6mo agoWhy? As a user of these tools, I love the convenience factor of having one tool rather than wrangling dozens. It's why in the past I've used an IDE (JetBrains), a language created by the provider of the IDE (Kotlin), web framework created by the same people (ktor), etc.
- kelvinjps10 6mo agoI feel like a lot of people and companies wanted to automate the web, but most website's operators wouldn't let you and would block you. Now you put the name AI into and now you're allowed to do It.
- sarpdag 6mo agoI can't pick the effort for the tasks run on Claude Web. I have a feeling Claude is using low or medium effort on those tasks, and I observe clear quality differences with the task ran on my local claude code, which uses high effort.
- cestivan 6mo ago[dead]
- georaa 6mo ago[flagged]
- Steinmark 6mo ago[dead]
- nlawalker 6mo agoMake sure to see channels too, just shared here last week - Push events into a running session with channels: https://news.ycombinator.com/item?id=47448524 https://news.ycombinator.com/item?id=47448524
- nickphx 6mo agoWho cares? Why does the hype machine need to hype the most inane 'features' as if they are novel, useful, or relevant?
- delphic-frog 6mo agoThe pricing discussion is interesting but I think people are missing the bigger picture. Being able to schedule agents to run tasks on a cron is genuinely usefull for solo devs who can't justify hiring someone to handle repetitive maintainence work. I've been using AI agents for image processing stuff and the autonomous loop is where it works.