8 ms·
TinyTroupe, a new LLM-powered multiagent persona simulation Python library
- turing_complete 2y agoWritten by George Hotz?
- deleted 2y ago[deleted]
- xrd 2y agoI love jupyter notebooks. And, I'm amazed that a company like Microsoft would put a notebook front and center that starts off with a bunch of errors. Not a good look. I really think you can improve your AI marketing by a lot by creating compelling jupyter notebooks. Unsloth is a great example of the right way. https://github.com/microsoft/TinyTroupe/blob/main/examples/advertisement_for_tv.ipynb https://github.com/microsoft/TinyTroupe/blob/main/examples/a...
- highcountess 2y agoI am glad I was not the only one that was taken aback a bit by that. I am not one to be too critical about loose ends or roughness in things that are provided for free and my ability to contribute a change, but it is a bit surprising that Microsoft would now have QA on this considering it ties into the current image they are trying to build.
- uniqueuid 2y agoNeeds openai or azure APIs. I wonder if it's possible to just use any openapi-compatible local provider.
- uniqueuid 2y agoYup, looks like their azure api configuration is just a generic wrapper for openapi in which you can plug any endpoint url. Nice. https://github.com/microsoft/TinyTroupe/blob/7ae16568ad1c4dea1381eda9dc8e0c3ff812d7a2/tinytroupe/openai_utils.py#L356 https://github.com/microsoft/TinyTroupe/blob/7ae16568ad1c4de...
- deleted 2y ago[deleted]
- deleted 2y ago[deleted]
- simonw 2y agoIt looks like this defaults to GPT-4o: https://github.com/microsoft/TinyTroupe/blob/7ae16568ad1c4dea1381eda9dc8e0c3ff812d7a2/examples/config.ini#L17 https://github.com/microsoft/TinyTroupe/blob/7ae16568ad1c4de... If you're going to try this out I would strongly recommend running it against GPT-4o mini instead. Mini is 16x cheaper and I'm confident the results you'll get out of it won't be 1/16th as good for this kind of experiment.
- ttul 2y agoI suppose the Microsoft researchers default to 4o because the models are free in their environment…
- cen4 2y agoMight help GRRM finish his books.
- deleted 2y ago[deleted]
- simonw 2y agoHere's a quick way to start this running if you're using uv: cd /tmp git clone https://github.com/microsoft/tinytroupe cd tinytroupe OPENAI_API_KEY='your-key-here' uv run jupyter notebook I used this pattern because my OpenAI key is stashed in a LLM-managed JSON file: OPENAI_API_KEY="$(jq -r '.openai' "$(dirname "$(llm logs path)")/keys.json")" \ uv run jupyter notebook (Which inspired me to add a new LLM feature: llm keys get openai - https://github.com/simonw/llm/issues/623 https://github.com/simonw/llm/issues/623)
- dragonwriter 2y agoThis seems fundamentally unsuitable for its stated purpose, which is “understanding human behavior”. While it may, as it says, produce “convincing interactions”, there is no basis at all peesented for believing it produces an accurate model of human behavior, so using it to “understand human behavior” is at best willful self-deception, and probably, with a little effort at tweaking inputs to produce the desired results, most often when used by someone who presents it as “enlightening productivity and business scenarios” it will be an engine for simply manufacturing support for a pre-selected option. It is certainly easier and cheaper than exploring actual human interactions to understand human behavior, but then so is just using a magic 8-ball, which may be less convincing, but for all the evidence supporting this is just as accurate.
- potatoman22 2y agoI wonder how one could measure the how human-like the agents' opinions and interactions are? There's a ton of value in simulating preferences, but you're right that it's hard to know if the simulation is accurate. I have a hunch that, through sampling many AI "opinions," you can arrive at something like the wisdom of the crowd, but again, it's hard to validate.
- kaibee 2y agocw: i don't actually work in ML, i just read a lot. if someone who is a real expert can tell me if my assessment here is correct, please let me know. > I have a hunch that, through sampling many AI "opinions," you can arrive at something like the wisdom of the crowd, but again, it's hard to validate. That's what an AI model already is. Let's say you had 10 temperature sensors on a mountain and you logged their data at time T. If you take the average of those 10 readings, you get a 'wisdom of the crowds' from the temperature sensors, which you can model as an avg + std of your 10 real measurements. You can then sample 10 new points from the normal distribution defined by that avg + std. Cool for generating new similar data, but it doesn't really tell you anything you didn't already know. Trying to get 'wisdom of crowds' through repeated querying of the AI model is equivalent to sampling 10 new points at random from your distribution. You'll get values that are like your original distribution of true values (w/ some outliers) but there's probably a better way to get at what you're looking to extract from the model.
- oulipo 2y agoSo fucking sad that the use of AI is for... manipulating more humans into clicking on ads Go get a fucking life, do something for the climate and repairing our societies social fabric instead
- thoreaux 2y agoHow do I get the schizophrenic version of this
- GlomarGadaffi 2y agoFollowing
- deleted 2y ago[deleted]
- itishappy 2y agoHere's the punchline from the Product Brainstorming example, imagining new AI-driven features to add to Microsoft Word: > AI-driven context-aware assistant. Suggests writing styles or tones based on the document's purpose and user's past preferences, adapting to industry-specific jargon. > Smart template system. Learns from user's editing patterns to offer real-time suggestions for document structure and content. > Automatic formatting and structuring for documents. Learns from previous documents to suggest efficient layouts and ensure compliance with standards like architectural specifications. > Medical checker AI. Ensures compliance with healthcare regulations and checks for medical accuracy, such as verifying drug dosages and interactions. > AI for building codes and compliance checks. Flags potential issues and ensures document accuracy and confidentiality, particularly useful for architects. > Design checker AI for sustainable architecture. Includes a database of materials for sustainable and cost-effective architecture choices. Right, so what's missing in Word is an AI generated medical compliance check that tracks drug interactions for you and an AI architectural compliance and confidentiality... thing. Of course these are all followed by a note that says "drawbacks: None." Also, the penultimate line generated 7 examples but cut the output off at 6. The intermediate output isn't much better, generally restating the same thing over and over and appending "in medicine" or "in architecture." They quickly drop any context this discussion relates to word processors in favor of discussing how a generic industrial AI could help them. (Drug interactions in Word, my word.) Worth noting this is a Microsoft product generating ideas for a different Microsoft product. I hope they vetted this within their org. As a proof of concept, this looks interesting! As a potentially useful business insight tool this seems far out. I suppose this might explain some of Microsoft's recent product decisions... https://github.com/microsoft/TinyTroupe/blob/main/examples/product_brainstorming.ipynb https://github.com/microsoft/TinyTroupe/blob/main/examples/p...
- potatoman22 2y agoThat example is funny because 99% of doctors would not use Word to write their notes (and not because it doesn't have this hot new AI feature).
- A4ET8a8uTh0 2y agoHmm, now lets see if there is an effort anywhere to link it to a local llm.
- ajcp 2y agoJust provide it the localhost:port for your instance of LM Studio/text-generation-webui as the Azure OpenAI endpoint in the config. Should work fine, but going to confirm now. EDIT: Okay, this repo is a mess. They have "OpenAI" hardcoded in so many places that it literally makes this useless for working with Azure OpenAI Service OR any other openai style API. That wouldn't be terrible once you fiddled with the config IF they weren't importing the config BEFORE they set default values...
- A4ET8a8uTh0 2y agoYeah, I was gonna say, I ran into all sorts of issues, but I couldn't immediately tell if it is my weird config or something. I don't want to give up on it, because the idea is interesting, but I will admit I only have so much time to spend on this today.
- ajcp 2y agoI finally got it straightened out after I unborked their hardcoding and implemented my own non-OpenAI/Azure OpenAI Service client.
- potatoman22 2y agoCare to share the fork?
- ajcp 2y agoAbsolutely but you'll have to give me a couple hours. I initially created a custom LM Studio client, but I've gone off the deep-end and am going to implement the Ollama client.
- thegabriele 2y agoI envision a future where ads are llms targetized. Is this worst or better than what we have now?
- libertine 2y agoCould this be applied to mass propaganda and disinformation campaigns on social networks? Like not only generating and testing narratives but then even use it for agents to generate engagement. We've seen massive bot networks unchecked on X to help tilt election results, so probably this could be deployed there too.
- Jimmc414 2y ago> We've seen massive bot networks unchecked on X to help tilt election results, so probably this could be deployed there too. Do you have more details on this?
- libertine 2y agoThere are so many instances over the past year but some examples: [0]https://www.cyber.gc.ca/en/news-events/russian-state-sponsored-media-organization-leverages-ai-enhanced-meliorator-software-foreign-malign-influence-activity https://www.cyber.gc.ca/en/news-events/russian-state-sponsor... [1]https://www.justice.gov/opa/pr/justice-department-leads-efforts-among-federal-international-and-private-sector-partners https://www.justice.gov/opa/pr/justice-department-leads-effo...
- isaacremuant 2y agoLet me guess. These bot networks that influence elections are from your political adversaries. Never from the party you support or your government when they're in power. The election results tilting talk is tired and hypocritical.
- libertine 2y agoSo this is all delusions: > The Federal Bureau of Investigation (FBI) and Cyber National Mission Force (CNMF), in partnership with the Canadian Centre for Cyber Security (CCCS), the Netherlands General Intelligence and Security Service (AIVD), and Netherlands Military Intelligence and Security Service (MIVD), and the National Police of the Netherlands (DNP) (hereinafter referred to as the authoring organizations), seek to warn of a covert tool used for disinformation campaigns benefiting the Russian Government. Affiliates of RT (formerly Russia Today), a Russian state-sponsored media organization, used this tool to create fictitious online personas, representing a number of nationalities, to post content on a social media platform.[0] Nothing is happening? Maybe the true question is, why are you going out of your way to make an effort to dismiss and normalize election interference and the usage of AI and bot farms to spread misinformation? [0]https://www.cyber.gc.ca/en/news-events/russian-state-sponsored-media-organization-leverages-ai-enhanced-meliorator-software-foreign-malign-influence-activity https://www.cyber.gc.ca/en/news-events/russian-state-sponsor...
- czbond 2y agoThis is really cool. I can see quite a number of potential applications.
- GlomarGadaffi 2y agoActually Sburb.