In short
Leveraging AI Podcast Episode 128 Summary
Episode Overview
- Title: The strategy and use cases used by a company to implement AI successfully with Peter Gostev
- Host: Isar Meitis
- Guest: Peter Gostev, Director of Data at Moonpig
- Focus: Real-world steps for AI implementation in a large-scale e-commerce company, with insights into organizational structure and practical use cases of AI.
Key Themes and Concepts
- AI Implementation Strategy
- Center of Excellence: Moonpig established a central AI team to drive experimentation and collaboration across departments.
- Purpose: Create a structured approach to AI, moving from individual experiments to company-wide implementation.
- Team Composition: Primarily made up of engineers with AI expertise. The focus is on using AI APIs effectively rather than solving complex machine learning problems.
- Use Case Development
- Process:
- Identifying Use Cases: Emphasizes the need for alignment with stakeholder priorities and feasibility.
- Three Buckets of Opportunities:
- Using Existing Tools: Implementing AI in areas like HR, customer relationship management (CRM), and communication.
- Engineering-Dependent Projects: Initiatives requiring development but likely to yield positive results.
- Exploratory Innovations: High-risk ideas with uncertain outcomes, such as using image and audio models.
- Stakeholder Engagement
- Importance of Collaboration: Successful AI projects require engaged stakeholders who are willing to contribute time and resources.
- Feedback Loop: Continuous communication with stakeholders is crucial for project viability.
- Training and Adoption
- Training Initiatives:
- Drop-In Sessions: Optional training sessions for employees to learn and experiment with AI tools.
- Hands-On Workshops: Sessions to build custom AI applications and increase familiarity with the technology.
- Cultivation of AI Champions: Encouraging individuals within departments to explore AI solutions helps scale initiatives.
- Practical Use Cases at Moonpig
- Customer Interaction Analysis: Leveraging AI to analyze chat and customer service interactions for sentiment and performance metrics.
- Automation of Routine Queries: Creating systems to automate responses for common customer inquiries (e.g., order status).
- Dynamic Product Tagging: Using AI to enhance e-commerce tagging and categorization, improving searchability and relevance.
- Future Directions
- Exploration of New Features: Ongoing experimentation with new AI tools and integrating them into customer experiences to enhance capabilities.
- Portfolio Management: Balancing high-impact projects with smaller, innovative initiatives to foster growth.
Key Takeaways
- Iterative Experimentation: The importance of piloting small projects and refining them based on feedback.
- Human-AI Partnership: AI should alleviate mundane tasks rather than replace human judgment and decision-making.
- Resource Allocation: Investing time in training and empowering employees can lead to sustainable AI integration.
Conclusion Isar Meitis and Peter Gostev provide a comprehensive look at how AI can be effectively implemented within organizations. By establishing a structured approach, engaging stakeholders, and focusing on practical use cases, businesses can harness AI's potential while ensuring ethical practices.
Additional Resources
- Moonpig AI Course: [AI Course](https://multiplai.ai/ai-course/)
- YouTube Full Episodes: [Multiplai AI YouTube Channel](https://www.youtube.com/@Multiplai_AI/)
- Connect with Isar Meitis: [LinkedIn](https://www.linkedin.com/in/isarmeitis/)
- Live Sessions and Newsletter: [Multiplai Events](https://services.multiplai.ai/events)
Feel free to subscribe to the 'Leveraging AI' podcast for more insights into the ethical and practical applications of AI in business!
Written by AI. May contain mistakes. Listen to the episode to check what was said.
Transcript
Automatic transcript. May contain errors.0:00Hello and welcome to another live episode of the Leveraging Ad Podcast. We have a really special episode for you today. Special number one, because I'm in Florida and there's a hurricane around us. And it's still pretty far, but I really, really hope I'm praying for the hurricane gods to keep my power up and running so we can finish this live session. But the other reason why this episode is special is maybe the two questions that I get asked the most. The first one is, how do I get started with AI? But since a lot of companies are kind of already gotten started, the second question that is very close second is now we kind of getting started.
0:42We have a few use cases, mostly specific individuals are doing like their own thing, but we don't know how to go from that to a company wide implementation of the process of how to implement AI in a more structured way as a company. I'm happy to tell you that's exactly going to be the topic and the focus of our episode today. Our guest today, Peter Gostev, is the head of AI at Moonpig. Those of you who don't know Moonpig means you're not in the UK. If you're in the UK, you would know them. So Moonpig is a really large company. They're actually a publicly traded company in the UK. And they sell personalized gift cards and personalized gifts.
1:22They obviously sell them online, but they deliver them in the real world, which means they're a very interesting mix of an online business combined with a real world business that actually has inventory and shipping and supply chain and customer service and everything that comes with living in these two universes, which makes Peter the perfect person to have this conversation with. So we are going to be discussing both strategy, how do we actually think about these concepts of AI implementations from a company-wide perspective, which is from his seat, that's what he's paid to do. So hopefully he knows what he's doing.
2:00No, he does. And you'll see that in a few minutes. But then also we'll dive into a few specific use cases that are practical ways that they're using AI right now at Moonpig. And so I'm really excited to welcome Peter Gustav to the show. Peter, welcome to Leveraging AI.
2:21In the next few years, AI technology will change our world dramatically. Whether you are a business executive trying to catapult your business forward, or just somebody who refuses to be left behind and want to advance your career, this is the show for you. I'm your host, Isar Maitis, a serial entrepreneur and an AI enthusiast. You'll hear invaluable practical tips from innovative business leaders, AI practitioners, and some of the brightest AI minds in our world today on how you can leverage AI in ethical ways to advance your career and grow your business.
3:02Lin, thanks a lot for inviting me. I'm looking forward to the conversation. Yeah, same here. It's very rarely that I get to talk to somebody who's running AI in that size of company? Like I talked to a lot of AI experts who are either consultants like me or work in smaller businesses. You really have a unique perspective because you're working on a sizable company. And again, I said in the beginning, it's a unique company because you're both in the digital and real universes. Tell us a little bit about from, again, a 30 ,000 foot perspective. How did you even get started with the process? What were the things that were done from a strategic perspective in Moonpig to even set the ground to enable the stuff we're going to talk about later?
3:48Yeah, sure. So I joined the company about, I think it's about nine months ago now. So we've certainly done things before I joined as well. And I think the recognition was that we probably needed more concerted effort around it where we needed to, apart from just general, please go try things and experiment, we needed to have some center of gravity where those sort of experiments could get going. And the hard thing about Generative Fire specifically is that people don't really have intuition for what works and what doesn't work. And that's why I think center of gravity is quite useful and that we've got a bit more space to experiment, to try things out, build out patterns for kind of things that work.
4:32And that was really the idea behind the team. We've got a central team. It's a fairly small team where we really experiment, we try things out, and then we partner with different other parts of the business to really help them also deliver. So we've got mixed mode. In some areas, we build new projects directly. We test ideas out and implement them. And in some other cases where we've got already a strong engineering team with some problems they couldn't solve maybe or couldn't easily solve on their backlog, we would partner with them and help them deliver better. So there's a kind of a bit of a central unit where we experiment.
5:13But still, I think what's important is that I've seen a lot of teams, maybe in my previous roles, in large organizations, where because it's so hard to innovate in the existing teams, they quite often would take innovation outside of those teams. It does help for some time, but then you can't actually deliver anything. So I'm a big fan of maybe experimenting a little bit, but you have to really get it to the teams who would actually build it and own it. So that's really the key idea behind the model. Yeah, I want to ask a follow-up question. But first of all, I want to thank all the people who are joining us live, both on LinkedIn and on the Zoom call.
5:54So I really appreciate all the people that are here. Feel free to introduce yourself, say where you're from, share a little bit about your company, share your LinkedIn link. If you're not on the LinkedIn side, if you're there, obviously people can see who you are. And if you have any questions, please feel free to write them in the chat. I'm monitoring both what's happening on LinkedIn and on Zoom. And if you want to chime in and be a part of the conversation, please go ahead and do that. So my first follow-up question to what you said is, in a quick summary, you basically created this center of excellence for AI where people have both the time, because they don't have a full-time day job to do other stuff, as well as the resources, meaning some kind of a Zoom or a playground and access to licenses, and you're experimenting with different ideas.
6:41My first question is, who are the people? Which department they come from? How did you pick them? Because one of the first things I do when I come into businesses for consulting and the things that I teach in my courses is to build an AI committee, which is basically what you have created. Who are the people on the committee? How do you pick them? What would be your recommendations to how to start a process like this? The team that I have is a very engineering team. So when I look for people to join the team, the way I think about it is that they should be engineers who then learned to do AI. But to be honest, there's definitely more than one way of approaching this problem.
7:29And I've definitely seen some very strong people who don't have AI background. They come in from research perspective or product perspective, and they just learn about how to use it. And then they apply it very well for what they need to do. Then I think probably the most popular path is to go from data science kind of background, machine learning background, and get into generative AI. And I think this is an interesting path that I think probably it is very popular because it's the closest in terms of the kind of the technology because it is machine learning, it is AI. But it's interesting that the mode of operating, I think, is quite different.
8:07That what we want to do with Genesify a lot of the time is, to be honest, just use the API and do the easy thing, not overcomplicated. And I think what data scientists, data engineers, or rather the machine learning engineers are used to doing is really to solve very difficult problems, prepare the data, train the model, and so on. And not that I'm against that skill set. I think it's actually really critical that we still not use the people who are very skilled at doing the hard thing. We just make them use the API. So I think it's actually important that we give those people the tasks that they can actually do really well.
8:46But I can also imagine a lot of people moving into the AI engineer role and actually not being as happy with that role because they actually just build software and use the API. But I think what's really important is that I think we have not really reached any level of maturity using the APIs. And certainly not at Moonpik yet. I think we've got a hundred things that we could do. And I could imagine it's probably the same at most other organizations. organizations. And I think what we should focus on is just basically use the API and find good ways of using the API to just in a very reductionist sense.
9:25Yeah, great answer. I'll add my two cents and then I have a follow-up questions when you said you have a hundred things that you could possibly do. My two cents to add is that even if you don't have technical people, there's a lot of stuff that you can do without having any technical skills. Like in most of the companies that I work with, we just either use the chat platforms as is or the automation tools that are built on top of them, custom GPTs or Cloud projects or Gemini Gems or whichever platform they want to use, or that in combination with external automation tools like Zapier and Make combined with different capabilities from the AI still allows you to do magical things.
10:08And the thing that I always recommend to companies in that committee is always to have people from each department participate in the committee, because then you get the inputs of what everybody needs, as well as you get different kinds of brainstorming, because it's different kinds of people. Finance people think differently than marketers or think differently than customer support people. And so you get the brainstorming that is a little better. And the thing that I think that is as critical is you get champions in each and every one of the departments to actually use the thing that the committee puts together.
10:39So beyond the techies, and if you have technical people in your company, yes, you want a couple of them in there because they can do a lot more stuff than you can do without those scales, but it's not a necessity and adding stuff on top of that is actually very helpful. So now to my follow-up question, that was a very long comment. I apologize. But you said a hundred things and yes, there's probably a very long list of things. So I have two follow-up questions one is how did you come up with the list of things that you can do and then what was the process to prioritize what are the things you actually want to do and actually maybe just to build on your comment in terms of the opportunities i would say for us we break it up into three buckets one is actually just a little bit like what you said in terms of using the tools and then we've got everyone from i know hr to comms to crm everyone using basically gpt pretty much we built our custom gpts in the areas where it makes sense and but we still experiment with new ideas and train people up and so on so that's definitely one bucket second bucket opportunities where we where we need to do some engineering but they're quite the things that we're pretty sure will work we just need to put some engineering behind it and then we just make it happen and then the third bucket is the more exploratory innovative bucket where we're not even sure it's going to work and for us quite often it is maybe image models where we want to see what is quite a visual business for us at Moonpeg so working with image models is quite interesting, but it's quite tricky as well.
12:21But not only that as well, maybe audio, video, all of these things come out and we want to experiment. So these are three buckets. In terms of how we come up with opportunities, I would say that there is a more kind of corporate answer that we've speaking to stakeholders and collecting ideas, evaluating ideas, prioritizing ideas in the kind of formal way. And that happens a bit. I would say what's probably pragmatically what's the most important thing for me is whether it is deliverable. And in the first instance, is it a thing that people care about? That's a given. But the second part, is it deliverable?
13:03And that has two parts to it. One is, does it technically work? And sometimes it just wouldn't work. and sometimes the issues where it's a, for example, something that we need to bring seven data sources together. And you know what? It's just not going to happen. We've got better things to do. So we just don't maybe not prioritize that one. But the other one, other category, which I think is really important is, does the stakeholder care? Would they actually put time behind this? And would they work with you? And would they partner with you, resolve issues, provide you extra documentation and feedback and so on?
13:41And I think that's probably the biggest thing because we can build the best theoretical prioritization model to calculate the value, calculate the uplift and so on. And then if the stakeholder says, you know what, I'm busy for the next 80 months, come back later, and you can't get past that point, then that's the biggest issue. In practice, I would say, apart from all the things that we all know that we should do, I would say the most important one is, can you actually work with the stakeholder to deliver it? No, I think that's a huge point. It goes back to the people you want involved in the process, right?
14:15In the committee, if you have the people who are the stakeholders, or at least their representatives, who can say, yes, this is a great idea. But like you said, I'm busy between now and March. So talk to me in April, then it's not a good idea to go down that path, even if on paper and the Excel file, the bottom right corner shows it's going to be a good ROI and so on. So I agree with you 100%. Something that in general is very obvious from everything that you're saying, and I agree with that from my personal experience working with multiple companies, is at the end of the day, the hardest part is the people part.
14:48The technology implementation is in most cases doable, sometimes easier, sometimes a little more complex, but it's doable. It's the people part that is usually the hardest. And you touched a little bit earlier when you said about training the people on how to build GPTs and how to do that. How do you do that? How do you go into what kind of training is delivered to employees? At what frequency? By what people? On what platforms? Do you evaluate them on that? What's the actual overall training process on a company-wide deployment of AI? I'm not going to claim that I've cracked it for sure, but we are trying.
15:28It's certainly not a soft problem to a point. I think it is about the people. And I think we did a bit of a survey across our chat GPT users. And the biggest point of feedback was, I don't have enough time to try things out. And I don't have enough time to dig into it and invest time into it. And I think that will be true because I think maybe if you're listening to this and thinking to yourself, or maybe that's easy, I use it all the time. The big difference towards probably 90 % of people is that it's not easy. They don't really know all the details, all the nuances about how it works. And sometimes when I speak to people who haven't used it before, don't really know about it much, you realize how much tacit knowledge you built up just using these tools every day about what kind of things work, what kind of things don't work, how much data you can put into it, what kind of data is usable and so on.
16:24So that's basically, that is a hard problem. So I don't think there's like a perfect answer. But I can tell you what we are doing at MoonPix. So we've got about nearly 200 users of chat GPG. which is a bonus is a lot i must say i didn't expect that many people to be interested and this is not really we did not make anyone sign up this was pretty much one slack message a number of months ago people just obviously teams it gets spread word of mouth and so on but i deliberately did not want to market it too much because i didn't want people just to sign up because it's cool and they don't want to use it so that that was the first stage and then we had a just kind of optional drop-in sessions for people to learn about how to use it that was part of it then also we had team sessions so i would try and pick out some teams who maybe expressed interest or for some reason i knew maybe they they want a bit more help and i would have team sessions with them and in those team sessions we would build a custom gpt together just go through a little bit of a training about how that would work.
17:31And then we also have sometimes big sessions with maybe the whole department. And we recently did one for our senior leadership where we also got them to build a custom GPT, which is quite cool. And I think it's probably the more senior you get, the harder it is to find time to just sit down and actually do the thing that you've been reading about. But I would say the overarching lesson out of all of this is that you must do the thing. just you have to try it it's just there's no the thing to realize is that there is no possible way how without trying you could develop the intuition for what works what doesn't work like what the precise intuition and if you don't do that then how could you know what what use cases are good or not and it's not that someone is i don't know not talented enough or something like that.
18:23It's not humanly possible for you to just emerge with that intuition. And even now, I'm still struggling to find a way to pass that intuition to people who haven't tried it much, and it's quite difficult. So I'm hoping I'll get better at that. But one other way that we do it is, what I find is that we can have all the training sessions, and then people still go back to their day job, and they don't really then necessarily find time to do the thing for four hours so sometimes i would still pick out a few people where we can't do the use case and we'll just sit down for half a day and just get it over the line and they're just that a bit more effort makes all the difference but yeah i could probably do that 30 more times and still have more ideas but i ideally what i'd like to do is to build more of a network where it's people more and more people do that we don't quite have that yet so there's definitely more that we can do in that space I want to summarize some of the critical things you said, because you touched on a lot of really important points.
19:25And again, I do this. I've been teaching AI courses to businesses, like private courses where companies hire me to do the things that you just said. Since April of last year, I taught hundreds of companies. And so I want to touch on some of the points that are very critical. The first and most important point is you have to block time on people's calendars to do this. And when I say block time is not give them a task on Slack, is actually have a meeting when they're going to show up and their department is going to show up and their boss is going to show up and they're going to be there for an hour, two hours, half a day, two days.
19:57That is a condensed time where you focus on learning how to use these AI tools. There is nothing that gets you more benefit than doing that, because as you said, people have day jobs. And even though people know this is important and they want to do this and they know the company is saying strategically, we've got to move in that direction, there's always more urgent stuff that happens. So this is the number one most important thing. The number two is what you said, that people have to experiment not with the concepts, but with the actual use cases that they're going to use. So the most effective thing that you said and that I do with companies is these are like mini hackathons.
20:34Here's the use case we're going to work on. Let's work on it for an hour, two hours, a day. It doesn't matter. But everybody sits in that room and working in a group or individuals, and we're going to solve this problem together using AI tools. And that does two things. First of all, it forces people to get their hands dirty. And like you're saying, the concepts are awesome. And you can listen to this podcast, which is fantastic, and follow people on YouTube and on LinkedIn and on TikTok. Wherever it is you follow people is great because it's going to give you a lot of ideas. but eventually you have your data, your company, your limitations, your co-partners, your processes, your licenses, your unique solution is going to be different than most of the stuff that you see by other people.
21:18These other people give you great ideas and starting points and cheats to jump through a few hoops, but you will have to figure it out on your own. And the only way to do this is to actually try out. And if you're going to try out for five minutes between meetings, it's not going to work. And then you're going to say, this is all bullshit and you're going to leave it. And that's going to be a big loss for you and your company. And you have to do these two things like free time for somebody who knows with a bunch of people who doesn't know to sit together after some initial training. So they have the basic concepts in place.
21:48What are the tools? What's prompting? What data you should not put in there? Like all these things needs to be there, but then dive into the actual hands-on, let's build this thing that solves this problem and let's invest the amount of time that's required because then people get the understanding of how it works, they understand what doesn't work. And then they also have something that actually does work in the end. And maybe it doesn't work perfect, but it works 70 % and they have enough data and knowledge and scale and excitement to go the other 30 % and then they're going to start using it.
22:16So great points. I want to move into the practical tactical side of things because I think we covered a lot on the strategy side. What are some of the use cases that MoonPig are currently using AI for in an impactful way that it's actually generating great results and people are excited about? Yeah. So one of the best categories that I would recommend to any business and certainly applies to us is basically looking at your unstructured data and seeing what you can do with that and the what i mean by that is any customer interactions any customer transcripts chat transcripts for example you could analyze them and say was this a good conversation or bad conversation what were the entities mentioned in this conversation was the did the agent do a good job of handling this and i know if they had a bad sentiment what were they talking about what which specific bit they're related on.
23:17Was this like a bug on a website, for example? And there are many things like that that you could just extract from the conversation. And I think most businesses, the way it worked for us otherwise is that maybe the agent raises it with their manager and then the manager raises with the developer team and hopefully get closed off, but it probably wouldn't happen. Then one of my favorite things is normally you get an MPS score or whatever feedback method you're using and maybe you get one percent of your conversations to get an MPS score whatever it might be you know I think I don't know what it is for us but it's like low digit numbers and but now you could get put it through LLMs and basically get the model to estimate what the MPS would have been for that conversation and it's not something we've deployed yet but through just light experimentation it seems to work really well so the kind of calibration seems pretty good so imagine now suddenly you go from maybe getting one percent coverage to 100 of coverage and by the way you can also do it fairly instantly okay maybe pragmatically next day now you can have a dashboard and you can have different correlations about why do you get that kind of i don't know why did you get that score you can dig into that a lot more.
24:40Maybe people talking about that specific problem. And so in terms of what we've deployed, we've done bits of it. The reason why we haven't done it instantly is actually data pipelines. And just we do have some data that we're pulling through, but not all of it. And the data pipeline is complicated. And anyway, real things like that that make it hard to actually deploy things in real life. So we're definitely going to push on that more. But in terms of high level experimentation that works really well. And in terms of... I want to pause you just for one second, because you touched on a few very important points.
25:17One of the maybe most magical things that these LLMs can do is qualitative data analysis at scale, which is something that was not possible, period, stop, before that. Because the only way to analyze large unstructured data like transcription of calls was to have people, lots and lots of people to go and listen to the calls, take notes, try to calibrate it in a way that they're all giving it the same level of score and whatever, which is impossible. If you have lots of data coming in and you have 50 people analyzing it somewhere in India where the math would make sense somehow from a cost perspective, how do you train those 50 people to give exactly the same scores?
26:00It's just very hard to do. And the only way to do this was to hire like a McKinsey who has this army of people to do a one-time project for you. Like it was impossible to do this ongoing. When I was running my travel company, I was running a hundred million dollar travel company. We had a call center to do customer service, to do outbound sales, to do inbound sales, like all of that. And what we had to do is there were people, full-time job people would listen to calls randomly because they can't listen to all the calls. And when there was something interesting happening, they would pick it up and we'll make it into a training session.
26:31And that's exactly your 1%, right? And now you could do this for 100 % of customer communication. It doesn't matter whether communication is verbal over the phone or on a chat or through email or through a third party tool like G2 software, if you're a software company, or it doesn't matter what the data source is you can take all the communication you have with customers as well as prospects and learn from it at scale. I want to ask an interesting follow-up question because you said you're already experimenting with this and I'm experimenting with this. So I just want to, I want to make it a little more tactical.
27:12How are you, even if you haven't deployed it and you just started playing with it, how are you practically planning to do this? Meaning, are you sending the data to LLMs and then putting it back into some kind of a database? Are you using a third-party tool like Intercom or something like that to do some of the work? What's the practicality of what you said? The likely pattern for us is that we'll get data out of a tool that we're using for customer service, then putting it probably in Snowflake, so just via ETL process. And in there, we've got choices with snowflake they've actually got some inbuilt lms now so you can do the analysis directly there so we we just got access we're experimenting with that another option was just to pull data output it through lms via whatever open air whatever ai tool you're using and then there's there is a question of how we exactly make that surface that to the end users one pattern we've got is to have a slugbot and basically just make that visible or just have passive updates via slugbot for now so it's not going to be like a chat to a data kind of use case but you just get updates and people subscribe to the channel and so on and something i want to explore as well whether we can make it more available as well to the managers directly for example in customer service centers so they it'll be more powerful if they've got access rather than people in some other departments have access.
28:42So there are probably a few things that we can do there. But yeah, we need to sort out a few of those steps in terms of pulling data in a way that makes sense. Things like that are hard. Yeah, it's a funny thing. AI doesn't really make that easy. Like you still need to do the data pipelines and so on and visualization and deployment. So there's one bit of it is a lot easier and then or it at least became from not possible at all to possible. but other things are pretty hard as well. Yeah, two cents to that and then I'll let you continue because I paused you in the middle of a sentence. This is a really important point for people but two things that I want to add quickly.
29:23One, you can test it at small scale on your own without any developers very quickly just using tools like Make or Zapier or NA10 that can grab any transcription that shows up on whatever platform you're using to transcribe it. And by the way, if you're not recording your interactions with your clients, please start recording them. Even if you're not going to use them now, at least six months from now, you'll have historical information to work with. So start recording everything. But you can literally grab the transcription of the recording from existing tools that either Zoom itself or Teams or phone recorders or whatever it is that you're using and run it through an automation tool with a predefined prompt through a chat GPT or a cloud or whatever and get whatever summary you want.
Read the full transcript
30:09And then you can run a different automation that every time there's 50 of these, go through those summaries and look for similarities and look for insights and so on. And that you can set up right now, today with zero developers, with one day of effort of tailoring everything together. So it's not going to be perfect, but it can be an amazing starting point at almost no investment. And so there's ways to solve this very, very quickly as a prototype in order to see what kind of benefits you can gain from that now let's go back to your process yeah and actually just to add on the prototyping the way i actually start every prototype is in just chat gpt to be honest not it doesn't work for every single thing but it's pretty much build a custom gpt and see how well it works and quite often the normal cycle for other project is typically i know you have a meeting with the stakeholders maybe then you do some prioritization then i don't know you get developer time maybe you build a prototype and so on and quite often i have a cycle is i have a meeting with the stakeholder in the morning and then afternoon i send them a custom gpt and we just see what do you think and then it doesn't mean that it instantly gets converted but at least we are having a real conversation about something rather than about hypotheticals about maybe this will work, maybe this wouldn't work.
31:32Yeah, and then there are other things that, yeah, but OpenAI Playground is great, so we can do that, which actually does have some extra features, for example, for structured outputs and JSON mode and so on, which is just a little bit tighter than what you can do in custom GPGs. So that's helpful. Yeah, so prototyping is super easy, and it's just worth always just do the prototype before just doing any prioritization sessions and so on. So then in terms of the use cases, so yeah, the field of just looking at unstructured data is incredibly rich. So we just really spend time on that. And it could be, so unstructured data includes things like, yeah, customer conversations, your internal documentations, that is a rich source.
32:18So if you've got already documents written, charter documents applications are quite nice and custom GPT is easier for that. But if you need something bigger, more robust, I think a Slack bot deployment with the documentation behind it are quite good for like HR policies or something like that. Then something we did as well is looking at all of our product descriptions and images and putting them through LLMs or vision models specifically of the LLMs. and then we improve the way we tag the products for e-commerce. And it's something that we probably spend, I don't know,$300 on in API costs or something like that, maybe$500.
33:06And even if the uplift is next to zero, the fact that we can just do that and put in the extra tax without having to go through a big project of doing that, It just makes it so much nicer. And the nice thing about it as well is that if you're halfway through, you decide actually you're going to change your tagging strategy, like you want more tags or fewer tags, you can just do that again. Just run it again with a different prompt and just test it. And you can swap between different tags. And so that makes it a lot more flexible. You're not locked in into some decision that you made earlier. And we do have different approaches with different brands, for example.
33:45So we can also test it and we certainly see an uplift depending on where we started from. Something we haven't done too much yet, but I'm keen at looking at all of the data we've got, such as maybe the customer journeys and looking at the logs of customer journeys. See whether we can put that through the lens as well and get them classified in a way that's probably very hard to do statistically. but maybe if you kind of reasoned about it, that like you and I looking at the logs, we can say, oh, you know what? This customer really struggled to look for, to find the product, but maybe statistically it'd be hard to say.
34:21So that's probably another area, but we haven't done that probably yet. I want to pause you for one second because on that particular topic, I've actually done something as a volunteer project. So there's a huge issue right now with antisemitism that's happening on college campuses in the US. And there's several different groups who are collecting that data from multiple sources, but they did not know what to do with the data. So they had data from actual students submitting it. They had data from news. They had data from the universities themselves that collected, like there's multiple sources all in open format.
34:54There was no one form that everybody was using for what to look for. And one of the big problems was exactly what you said, like how do you categorize it? Like how do you define the categories that then you can go and look for and so on? And we literally did just that. We put everything in one CSV file and it has, I don't know, like thousands of rows. And we told the large language models to tell us how it would be the best way to categorize it so we can have something actionable about it. And it gave us several different options on how to categorize it. We picked one and we finessed it a little bit, but most of it was as it came from the large language models.
35:30And then we use that to categorize it. So it's even the first step, like sometimes you have all that data, like you're saying, you don't even know what you're looking for. There's so much information there. What can I learn from it that is beneficial for my goal? In my particular case was to learn about antisemitic events. In your case, it's how do people engage with our products? What are they looking for? What are they not finding? It doesn't matter. But so the AI is really good at finding patterns. And so it can find the patterns for you and can recommend things aligned with your needs, if you'll define it to the LLM.
36:10So yes, there's even the first step in the process you can solve that was, like you said, almost impossible to do before. Did you have any problems with like models not wanting to engage with that kind of language? so the way i'm doing this and maybe that's why i didn't have a problem but it's i'm bringing all the data to a csv file i'm opening it in google sheets and i actually have a code in google sheets that allows me to bring multiple large language models into the google sheets itself and it's just running on the apis in the background i'm using a tool called open router so open router For those of you who don't know the tool, it's like a bridge to almost any large language model API out there.
36:59So you get one API call, but then you can pick which model you're going to use. So the way I'm using it in Google Sheets, I have four or five different columns with the same prompt running on four or five different models. And then I can see which one is working best for that particular use case. And then if there's more than one that's working good for that use case, I go for the cheaper one. Some of them cost three cents for a million tokens. Some of them cost$70 for a million tokens. The spread is pretty big. Any one of them is still cheaper than giving a person to try to do the same work. But if you can save two orders of magnitude, why not do that?
37:34So that's how I use it. I never had issues with the API not working because it didn't want to deal with that data. Yeah, yeah. Okay, okay. Interesting. Yeah, that's actually, in terms of doing the categorization initially, that was actually one of the use cases that worked pretty well for me. with the latest OpenAI model, the O1 preview. Because I have tried to get it to do the categorization initially. And maybe for that kind of use case that you're describing with the Angus images, maybe it can understand it well enough so it can work out the categories. For us, I found that quite often the things we would care about is not something a model can guess.
38:19But for example, I would care, we are doing some debugging on the tool that detects, that basically helps customer service agents just tweak the messages automatically. And we want to see what was the original message that they got and what was the one that they sent and basically analyze the difference. And the things that I would care about detecting there is maybe not what the model cares about. So I don't really care about formats. that's more personal preference, but I would care about whether it removed Moonpick branding, for example, because that was a recurring issue. Everyone attested it.
38:57But it was interesting. I would say the O1 preview actually did a really good job on that. So I was using it quite a lot. And I was saying, yeah, here's all the data. Go now, categorize it and make sure it's like Missy and all of that. So it did do a good job. Awesome. Awesome. So let's do a quick recap of use cases. We talked about, in general, unstructured data analysis, wherever it comes from. We talked about customer service. We talked about categorization of specific types of data, doesn't matter where it comes from, so you can do more with it. What other use cases are you guys using it for right now?
39:39yeah so then the so within customer service there are quite a few like smaller things that that we are looking at so one i think a lot of the time when people look at customer service they say oh we can build a chatbot that just kind of automates everything and i think that's fine i think we do as well and a lot of the questions we get where's my order kind of question and to be honest it's not interesting for a human to respond to they don't really add any values it's not it's not a fun job so like that part is more or less automated and then when we looked at the processes of what agents are doing as well we found that a lot of the time they would select like a template that they would respond with to a specific issue but then they would just spend a lot of time like changing it basically to insert the customer's name or insert the issue that But things like that, it's not fun to do.
40:37It's not really adding any value. So we came up with an idea of basically having a little application on the site where we can basically take in the context of the conversation so far, embed extra rules behind the scenes where we are basically saying, if it's a chat channel, then you should format the message in a chat way. If it's a channel that's like email channel, it should be email channel. that kind of stuff took a lot of time as well just no value added they're just reformatting your message and the important thing is the important decision we made there is not to be too ambitious in terms of how much we want to how much thinking we're going to give to the model and actually more or less maybe it's a bit harder for new starters but the agents know how to solve issues they know they can think about But if it should give refund or we shouldn't give a refund, or if so, then how much?
41:39And there's some discretion there. Sometimes they go to the system and be like, well, we probably wouldn't give a refund normally, but this is a really good customer for us, so we should give it, and so on. So we basically decided we're not going to give the model any power like that, and we are not going to rely on it for judgment. Judgment still stays with the humans. and actually the change that what it meant for us in terms of the scope of the project and the complexity of the project was probably like 10x of we had to do so much more work to get it to perform even understand when you're supposed to give a refund it's actually quite a hard question because you need to know what kind of policies we've got for giving a refund we need to know what kind of information we even need to know before we can give a refund then we need to have a judgment of i don't know send us a picture of your gift being like destroyed there's so many things and this is just for one set of use cases and to be honest it's so much easier for us to say you know what for now as a first version let the agents who are humans who are good at their jobs just do the hard part and we'll just do the thing that they don't like doing and we'll just automate that and so far it's it's only been live for a little bit but the feedback's really good it's doing a good job there's still like some careful work we need to do to make sure it works well but i really like that that i think we we made something that we could actually deliver that works well that we don't need to maintain constantly in terms of like knowledge base and so on, but we're still living.
43:21I hope we're making the job more fun, more engaging, that they're not doing the tedious parts of their work and actually just helping customers. And I personally like that, but that's a use case I feel good about. Yeah, I want to, Cassie Kozikov, who used to be the chief strategy officer for Google, so a very smart woman. She's not there anymore. She left, she's now doing her own thing. But she had a very famous lecture that she has a great YouTube channel, by the way. look up Cassie Kazikov, anybody who's listening on her YouTube. She talks a lot about how she brings really complex, advanced AI concepts into monkeys like me to understand.
43:59So it's a great channel to follow. But she has a very famous lecture about thinking versus thunking. And I actually got to see her live talk about this about a year ago. And she talks about some of our job is thinking, which what you're saying is like your human judgment or whether that makes sense or does make sense. But then there's the thunking part, which is everything other than thinking, right? it's a word she made up, which is how do I now write the answer? How do I format the answer? What format, like which platform do I need to put this? All of that, nobody likes to do. It's just overhead that we have to do, or we did so far.
44:33And now that could be completely eliminated or mostly eliminated or assisted by AI. And so actually the direction you're describing is awesome because it's a significantly lower investment to develop the solution. You're keeping the employees versus the Klarna case, we said, oh, we can let go of 700 customer service employees because the bot now does it. And you're letting people be happier with their jobs because they're really focusing on helping people and making the right thing versus having to deal with crap they don't want to deal with. So I agree with you 100%. It's an awesome solution.
45:09And I think with the Klarna-like approach, we did consider doing a project along those lines. And the thing you realize is that for us to actually implement something, I guess, the key missing element is actually the context that the model should have. It's basically the data. But I don't mean, when I say data, I don't mean it in the sense that, I don't know, we didn't collect enough data or something like that. What I mean is it's more the LLM documentation that you need to write in a way that's not just good for humans, that is written for LLM. To give you an example, we have a policy. It's just a little table which just shows when you're supposed to do what.
45:55Give a refund. If it's delayed by more than a certain number of days, then you can give a refund, that kind of thing. if you just to give if you were to take this as a document to give this to the model it would just it would make no sense to that model you have to provide so much context about what is it that that is happening you need to describe all of your processes and by the way you have to do it perfectly because they cannot go and speak to their manager their life right what a human agent can do. So there's no real room for tolerance. And the threshold for actually getting this right is so much work, so much documentation you have to write, perfect maintenance.
46:38And I think, I don't know what CloudNet is doing, but I imagine they would be doing is something like they probably have a big operation to actually maintain the documentation and make sure that It's up to date. It's correct. It works well. And yeah, I think that is a promising direction as well. I think you can go down that route, but it's not free. You can't just like plug in the model and it just works. So I think maybe for operation that size, it might make sense. I think for us, it will be too much, too big of a project. And to be honest, there's so many other ideas I want to explore. I don't really want to spend the next eight months writing documentation for chatbots.
47:18so you you bring a great point and i actually want to follow up on that one is that it's an roi game right yes you can can you do this what clarna did probably yes it's going to cost you x number of dollars and x number of months to develop this and is it worth doing and let's say even if it is worth doing let's say it is going to yield a positive roi within one year but what what are you not doing at that time? Because you're not doing other projects and other initiatives because you're investing all your eggs in that basket. So my question to you is how do you currently really pick those projects?
47:55You gave us a general idea, but you said you didn't want to do this one. You have other ones. What are the things that you're looking into right now, as an example? And if it's sensitive, you don't have to say that. At least tell us the concepts, because I think that would be very interesting for people to know. What are the things that somebody in your position at that size of company is looking into as far as implementing in the next six to 12 months. Yes. And I'll start with a bit of framing is that I want to have a portfolio of things. So I want to have big things that I'm delivering to make sure that there is impact.
48:26And then I want to have little things that we can go on day to day and deliver. So part of the little things are things like running training sessions and picking up with specific teams and just helping them all kind of thing there are there's a really nice category of use cases where it's the teams just need like a little bit of inspiration a little bit of push and then actually they could just go and explore those use cases so i don't think i can say some specific things we're doing just because it links to other things and so on but generally it's where they have a some specific problem that and basically can't launch the product or it makes the user experience so much worse.
49:09And then what I try and do is to be in the places where those kinds of problems are discussed. And then I could basically say, you know what? Actually, if you did use LLMs, you can probably do that a lot easier. And quite often, I think people, A, just wouldn't know how it works. They don't have the right intuition for picking those problems up. And my job is to have that intuition and to pick up the problems and then have the intuition to solve them. And then they just never actually tried it. And there's a little bit of a barrier of, yeah, even if you have the intuition, but to literally see like where do you click?
49:50What does the API structure look like? Oh, actually, yeah, you have to actually define the schema here. There's still like a bit of friction there that just developers wouldn't necessarily know out of the box. But once they have the kind of ergonomics of using it, it's actually they just go and build it. And I had probably three examples like that already where my involvement was probably like, I would say a couple of days, and then the teams just went and built something. And yeah, these kinds of projects are my favorite because I can do like nearly unlimited number of those. And then the teams just do their normal job and they can just do it a lot better.
50:34But yeah, and then I mentioned experimentation. So certainly we try and think about any new things that we could potentially do. Many new models come out and we just explore built little prototypes and experiment. And then hopefully you'll see some things come out in the next few months from Unpeg. But yeah, it's basically new customer features of this stuff that we just couldn't do the last six months and now we can. So that's the kind of things that we're experimenting with as well. That's awesome. I'll piggyback on the geeky side of things and going back to our very first point. And that would be a great way to have full closure.
51:13And then we can say thank you. But in the committee or like in maybe you have subcommittees for specific things, you want to have geeks as part of the team. because the geeks of the company who love playing with these tools, who find it exciting, who are like me when a new model comes out or a new tool they found will spend between 10 p.m. and 2 a.m. playing with this to try to figure it out can be exactly those people that can do the thing that Peter just described. If you will let them be the champions of these little projects, you will have significantly more successful projects across multiple departments just because you let people do it.
51:51And these people can become your champions within the departments to do the stuff that Peter cannot do all on his own because he's one person. And so if you have a few more people in the company who are actually enjoying doing this and they have that intuition and understand what can be done and you give them the freedom and you give them the quote unquote title of, okay, you are now the AI champion of the finance department and you are allowed to use these tools and you're not allowed to touch this kind of data, but now go do whatever you want with it, you will get amazing stuff out of this very quickly, exactly in those kinds of scenarios that Peter described.
52:29Peter, this was a fascinating conversation. I personally learned a lot. I'm sure that people are listening, learned a lot as well. If people want to follow you, work with you, learn more about your journey, what are the best ways to do that? I think probably the only way, the best way is on LinkedIn. So I I think that's probably the easiest way. You can connect with me, message me. I'll try to respond, but I sometimes don't check my messages for a little while, but I will get to you at some point. But yeah, this was a great conversation. Yeah, thank you so much for inviting me. I really enjoyed it.
53:05No, thank you. And I want to thank again, all the people who joined us live, who are here on Zoom and on LinkedIn Live. I know this was very valuable because this is a phenomenal conversation that is really critical to anybody who's trying to do this in AI. so again thank everybody thank you peter have an amazing rest
From the publisher
Join us for an exclusive, live discussion with Peter Gostev, Director of Data at Moonpig, as he takes you on a deep dive into the real-world steps behind AI implementation at a large-scale e-commerce company. Peter will share his experience navigating AI integration in a 500-person organization, from organizational structure to use case selection.
In this episode of Leveraging AI, Peter will reveal how Moonpig uses AI for automation, providing actionable insights every business leader can use. You’ll leave with the tools and knowledge to replicate Moonpig’s AI-driven success in your own company.
Peter is a leader in the AI space, driving innovative solutions at Moonpig, one of the UK’s largest online greeting card and gifting companies. His approach combines practical AI strategies with a clear focus on business results, making him a must-hear for any AI enthusiast or business leader looking to scale AI in their organization.
---
Join the next open AI course: https://multiplai.ai/ai-course/
About Leveraging AI
- The Ultimate AI Course for Business People: https://multiplai.ai/ai-course/
- YouTube Full Episodes: https://www.youtube.com/@Multiplai_AI/
- Connect with Isar Meitis: https://www.linkedin.com/in/isarmeitis/
- Join our Live Sessions, AI Hangouts and newsletter: https://services.multiplai.ai/events
If you’ve enjoyed or benefited from some of the insights of this episode, leave us a five-star review on your favorite podcast platform, and let us know what you learned, found helpful, or liked most about this show!



