UnionFaps
AIExplained from patreon
AIExplained patreon

SmartGPT Website Demo and Community Project

🕑 Added 2024-04-12 15:09:48 +0000 UTC

Comments

Pranav S

is it possible you can add claude 3.5 to model or allow for us to set to newest models?

Joshua Davis

I tried a few different times and I found that the output was poorer than if I simply used Claude 3 or even Gemini Pro 1.5... sorry. I'm happy to share my chat with you, maybe I'm doing something wrong.

Joshua Davis

I keep getting this error: Message limit is {{maxLength}} characters. You have entered {{valueLength}} characters.

Christopher Pollin

Wow! I had time to try it out today! I do a lot of data modeling for digital humanities and digital history. I was able to create (almost) a complete data model, just based on historical sources in plain text as input. I already had some sophisticated prompts, but this is the next stage. Right now, agent-based workflows are a hot topic, aren't they? Your smartGPT could even be scaled up by using multiple teams of agents based on smartGPT... what do you think? And today i listened to 4 professors at a panel discussion at my university about "ai and research". all they talked about was how to write papers and that students submit ai texts. at the same time, with the big context windows we already see and the next generation of llm and methods/tools liek smartGPT, you can already imagine that whole research processes could possibly be automated. crazy! :)

Robert Gomez-Reino

Thanks Philip, I meant stored in clear on your side. Just trying to assess what type of test we should or not perform. thanks!

Philip

Hey Robert, yes if you click save to template top right that will store modified system and researcher/resolver prompts! Trying to keep things fairly user-friendly, while still allowing customisation.

Philip

Thank you John! Keep us updated.

Robert Gomez-Reino

My feeling is that this agentic approaches with gpt4-like capabilities need to be balanced with speed. I want it to be fast enough (really few seconds) to be able to steer results and repeat things. We've put substantial effort in previous months in collaboration with CERN and the amount of work to align workers, resolvers, reviewers to any general (in our case engineering) task become increasingly difficult. You end up many times waiting for mins to get an answer not aligned with the project patterns or just wrong. I am now convinced that shorter (but much faster) interactions, even if 60% times only ok, is much more productive because you can steer, request corrections, etc. and thus have a realistic working flow. More agentic could still be ok for more restricted set of results you can tune for agents for IMHO. Otherwise also very interesting to keep researching on this flow, so when the next thing (GPT5?) comes we just need to plug it in :D

Robert Gomez-Reino

Philip, would our prompts be stored somewhere when testing this tool? :D We have our own tailor made version and I would love to check how your's improve our chain of thoughts arch. Our current one requires large system prompts (like 20k tokens or so), I see we don't have this as I assume you are providing your agents your own system prompts, do you think this would make it difficult to work providing this large context information in user prompts?

John Weisenfeld

Kudos Philip, Yannick, Josh. My use case comes from my teaching. I recently gave a two-question quiz (in a Microsoft Form) to my students to "give me something true (and non-trivial) about the Big Bang" and "give me something false (and non-trivial) about the Big Bang". I input a short rubric plus a student's response, and I'm impressed that SmartGPT can analyze student answers in-depth and give a total score which is consistent with the rubric. Well done!

Kol Tregaskes

Hi Philip, still not working, let me try other browsers. I use Brave, which can be temperamental sometimes.

Leander Maerkisch

Yes, I did! Thanks a lot. Was just extremely busy with a new product release until Friday & slept through the weekend. Will answer later today!

Philip

Thank you! And trust me, I am as ambitious for it as you are with these improvements, all while keeping in mind an average user who has never seen an API before.

Philip

Thank you Sean.

Philip

Thanks Leander! Hope you got my earlier email re: Alberto!

Machiel Reyneke

Also, it's crazy to think that we're at the vacuum tube stages of genAI - what will the microchip stage look like?

Machiel Reyneke

Thanks for sharing SmartGPT, Philip! I've tested it, and so far super-impressed :) Used Haiku as the assistant with Opus as the resolver and it worked well for a few typical product management use cases. It'll be interesting to see how these agentic workflows evolve - I would personally love to have a blend of Autogen (with custom agents, including a default setup like the SmartGPT "Assistant", "Researcher", "Resolver" pattern), alongside your implementation of N parallel assistants. I guess later the best would be if a model can classify the initial prompt in terms of difficulty and automatically set up all the necessary agents to do all the interactions, and let the harder steps be N parallel agents with resolvers ala SmartGPT.

Sean Gallagher

Philip, Relieved to hear you avoided a September demise. We’d all be immeasurably poorer without AI Explained and Insiders. You should be immensely proud, not just in explaining AI, but in shaping its future.

Leander Maerkisch

Wow, congrats on the launch of this new "product"! The folder structure to store similar conversations is already a UI improvement over the classic ChatGPT interface. Well done & excited to test the waters!

Philip

Thanks so much Dorian, excited for where this will go. I will stick with that phrasing then!

Philip

Any updates Kol? :) As pretty much my earliest subscriber on YT...

Dorian Iten

Fantastic, thank you so much for sharing this with us, Philip! By the way, "hype-free AI content" is a very compelling phrasing for me. 👍

Philip

You'd be surprised what use cases it is good at! Quite a few! So glad to hear Steve :)

SteveHaupt

I am impressed. It solves my default thinking problem with which I test all models consistently. No model so far could solve it, only the new gpt-4-turbo sometimes gets it right and ChatGPT with browsing and code interpreter is a bit more reliable but still often wrong. What is my default thinking problem: How will the time difference between Munich and Sydney change in 2024? Gives the answer in json format. Start with 01.01: time difference, continue with all dates where the value changes and end with 31.12: time difference. Its a difficult problem, bc Germany and Australia change time at different dates and in different directions.

Philip

Kol and Jon, we are getting error messages about incorrect API keys and insufficient balance with Anthropic, and 1 error for quota limit with OpenAI, so do try again at some point with those fixed! Would love your feedback.

Philip

Thank you so much Christopher, who knows how many are here because of you.

Jon Millward

I tried again without changing anything and it worked. I suspect it's a time-out issue, because when it doesn't work, I see "Initial GPT Answers (3 Asks, Model: gpt-4-turbo)" (Generating response), then blankness for 30 seconds or so, then no output and "Regenerate response". No red error message. It's like it gave a response when it actually didn't. Maybe it didn't have time to generate and evaluate three responses before it timed out. Perhaps there's some way to show the progress of the request in more detail? Anyway, I won't flag any more issues in here as it's not the right place. I'll keep playing around with different length requests. Thanks.

Philip

We will try to get to the bottom of it, looking now through our error messages. Did you re-submit the API keys? I checked it and working my end, any red error message up top?

Jon Millward

This is what I see too. It did work once but not since.

Kol Tregaskes

Same here, I'm using GPT-4-Turbo, added my API, asked it a question and now I'm stuck on "Initial GPT Answers (3 Asks, Model: gpt-4-turbo):"

Kol Tregaskes

Oh my. I've been waiting for this for a very long time. Thanks, Phillip.

Dane Wagenhoffer

Ooh didn't finish reading since I'm at work, I can definitely do that.

Philip

Working on that, first Gemini 1.5 but then those, yep. Feel free to help too!

Dane Wagenhoffer

Can you add Cohere or Mistral API support? I have pretty much switched all my scaled LLM use to cohere so it might be interesting to test

Philip

Just checked, working for me. If it persists, email via address in Info. Sometimes it times out for really long tasks, with many asks, test it on slightly shorter.

Philip

Haha

Stefan

yep, it's working great Philip!

Tony Coffman

{to be read in the most sarcastic tone one could possibly muster up} Philip, if you would simply stop doing research and just believe all the hype out there... All you need to do is take a "FREE 10-minute course on AI", "subscribe to ChatGPT", or "download an AMAZING and FREE open-source model from Hugging Face". You don't need to hire anybody. Just "use AI". Obviously, AI can "easily" simplify every aspect of your personal and professional life, as well as all your job functions. "Hire an AI agent" sit back and let the subscription dollars roll in!

Jon Millward

It's not outputting any responses for me despite giving it API keys for Claude and GPT, both with credit. I see the credits being used in Usage for both, but no output...hmmm. Will keep trying.

Norfuer

I'll definitely have to give it a looksee in that case. My only concern would be the, I imagine, ginormous API charge if I roll it out often enough. I do two articles a day, six days a week. Some of these things go up to 10k words. Not all, but some. I guess I'll start testing by throwing in a whopper and see how much it'll cost. Thanks for all your hard work, Philip. You really stand out among the crowd of AI content creators, all without having to publish everyday (on potential nothingburgers) and SHOCKING the audience with your titles!

Christopher Pollin

Thx! Your YouTube channel and Patreon are already the most recommended and quoted resource in all my workshops! :)

Philip

I mean, if you share it in the context of promoting the Patreon that sustains it, then I would be more than happy! For most situations that might be in the form of a visual demo from you that they can then play with if they sign up, but for key networkers you feel might really be able to spread the word then a simple link would be effective. So semi-private is how I would best phrase it (if I wanted 100k users, I would just put the link on my channel!). I trust your judgement, this post was more to thank my Patrons and get early feedback before we iterate in the coming weeks.

Philip

I have found it great for summaries. A few individual calls often miss a detail here and there, but rarely after going through the full SGPT process. Obviously might have to adapt the system prompts for your use case.

Norfuer

For work, I summarise (often lengthy and highly technical) scientific papers for social media. Been using a GPT for it since that launched, but I feel it's not quite as smart as it used to be. The outputs tend to be too vague and lack the meat of what I'd like to be a more in-depth but still comprehensible summary. I wonder if this is a valid use case for SmartGPT, to have it output layman-friendly but accurate summaries of often highly technical material.

Philip

Yeah, this system alone would get SOTA on reasoning, but multi-modality is where it gets really exciting.

Christopher Pollin

cool! i do a lot (really a lot) of genai & prompt engineering workshops for researchers (data modelling, creation, programming, thesis and proposal writing, etc.). having something with more reasoning power would be really cool to do some showcases. is it ok to share the link, or should it stay internal? and if so, how should it be cited :) In other words, is it in your interest to get it out into the world, if others show it at workshops? best, Christopher

Trenton Dambrowitz

Some very impressive first impressions on a variety of questions, very interesting to see how it works with nuanced tasks like creating repair plans from images. Nice Job Philip!

Nathaniel DiMemmo

I’ll play around a bit this weekend. Congrats on getting it published.

Adrian Ott

Can’t wait to try it out!


More Creators