examples: remove codified examples (#8267)

84a23144 · Parth Sareen · GitHub · 17fcdea6 · 17fcdea6 · 17fcdea6
Unverified Commit 84a23144 authored Jan 13, 2025 by Parth Sareen Committed by GitHub Jan 13, 2025
20 changed files
--- a/examples/modelfile-mario/readme.md
+++ b/examples/modelfile-mario/readme.md
-<img src="logo.png" alt="image of Italian plumber" height="200"/>
-
-# Example character: Mario
-
-This example shows how to create a basic character using Llama 3.2 as the base model.
-
-To run this example:
-
-1. Download the Modelfile
-2. `ollama pull llama3.2` to get the base model used in the model file.
-3. `ollama create NAME -f ./Modelfile`
-4. `ollama run NAME`
-
-Ask it some questions like "Who are you?" or "Is Peach in trouble again?"
-
-## Editing this file
-
-What the model file looks like:
-
-```
-FROM llama3.2
-PARAMETER temperature 1
-SYSTEM """
-You are Mario from Super Mario Bros, acting as an assistant.
-"""
-```
-
-What if you want to change its behaviour?
-
- Try changing the prompt
- Try changing the parameters [Docs](https://github.com/ollama/ollama/blob/main/docs/modelfile.md)
- Try changing the model (e.g. An uncensored model by `FROM wizard-vicuna` this is the wizard-vicuna uncensored model )
-
-Once the changes are made,
-
-1. `ollama create NAME -f ./Modelfile`
-2. `ollama run NAME`
-3. Iterate until you are happy with the results.
-
-Notes:
-
- This example is for research purposes only. There is no affiliation with any entity.
- When using an uncensored model, please be aware that it may generate offensive content.
--- a/examples/python-dockerit/Modelfile
+++ b/examples/python-dockerit/Modelfile
-FROM mistral
-SYSTEM """
-You are an experienced Devops engineer focused on docker. When given specifications for a particular need or application you know the best way to host that within a docker container. For instance if someone tells you they want an nginx server to host files located at /web you will answer as follows
-
---start
-FROM nginx:alpine
-COPY /myweb /usr/share/nginx/html
-EXPOSE 80
---end
-
-Notice that the answer you should give is just the contents of the dockerfile with no explanation and there are three dashes and the word start at the beginning and 3 dashes and the word end. The full output can be piped into a file and run as is. Here is another example. The user will ask to launch a Postgres server with a password of abc123. And the response should be
-
---start
-FROM postgres:latest
-ENV POSTGRES_PASSWORD=abc123
-EXPOSE 5432
---end
-
-Again it's just the contents of the dockerfile and nothing else.
-"""
--- a/examples/python-dockerit/README.md
+++ b/examples/python-dockerit/README.md
-# DockerIt
-
-DockerIt is a tool to help you build and run your application in a Docker container. It consists of a model that defines the system prompt and model weights to use, along with a python script to then build the container and run the image automatically.
-
-## Running the Example
-
-1. Ensure you have the `mattw/dockerit` model installed:
-
-   ```bash
-   ollama pull mattw/dockerit
-   ```
-
-2. Make sure Docker is running on your machine.
-
-3. Install the Python Requirements.
-
-   ```bash
-   pip install -r requirements.txt
-   ```
-
-4. Run the example:
-
-   ```bash
-   python dockerit.py "simple postgres server with admin password set to 123"
-   ```
-
-5. Enter the name you would like to use for your container image.
-
-## Caveats
-
-This is a simple example. It's assuming the Dockerfile content generated is going to work. In many cases, even with simple web servers, it fails when trying to copy files that don't exist. It's simply an example of what you could possibly do.
--- a/examples/python-dockerit/dockerit.py
+++ b/examples/python-dockerit/dockerit.py
-import requests, json, docker, io, sys
-inputDescription = " ".join(sys.argv[1:])
-imageName = input("Enter the name of the image: ")
-client = docker.from_env()
-s = requests.Session()
-output=""
-with s.post('http://localhost:11434/api/generate', json={'model': 'mattw/dockerit', 'prompt': inputDescription}, stream=True) as r:
-  for line in r.iter_lines():
-    if line:
-      j = json.loads(line)
-      if "response" in j:
-        output = output +j["response"]
-output = output[output.find("---start")+9:output.find("---end")-1]
-f = io.BytesIO(bytes(output, 'utf-8'))
-client.images.build(fileobj=f, tag=imageName)
-container = client.containers.run(imageName, detach=True)
-print("Container named", container.name, " started with id: ",container.id)
--- a/examples/python-dockerit/requirements.txt
+++ b/examples/python-dockerit/requirements.txt
-docker
\ No newline at end of file
--- a/examples/python-grounded-factuality-rag-check/README.md
+++ b/examples/python-grounded-factuality-rag-check/README.md
-# RAG Hallucination Checker using Bespoke-Minicheck
-
-This example allows the user to ask questions related to a document, which can be specified via an article url. Relevant chunks are retrieved from the document and given to `llama3.2` as context to answer the question. Then each sentence in the answer is checked against the retrieved chunks using `bespoke-minicheck` to ensure that the answer does not contain hallucinations.
-
-## Running the Example
-
-1. Ensure `all-minilm` (embedding) `llama3.2` (chat) and `bespoke-minicheck` (check) models installed:
-
-   ```bash
-   ollama pull all-minilm
-   ollama pull llama3.2
-   ollama pull bespoke-minicheck
-   ```
-
-2. Install the dependencies.
-
-   ```bash
-   pip install -r requirements.txt
-   ```
-
-3. Run the example:
-
-   ```bash
-   python main.py
-   ```
-
-## Expected Output
-
-```text
-Enter the URL of an article you want to chat with, or press Enter for default example:
-
-Loaded, chunked, and embedded text from https://www.theverge.com/2024/9/12/24242439/openai-o1-model-reasoning-strawberry-chatgpt.
-
-Enter your question or type quit: Who is the CEO of openai?
-
-Retrieved chunks:
-OpenAI is releasing a new model called o1 , the first in a planned series of “ reasoning ” models that have been trained to answer more complex questions , faster than a human can . It ’ s being released alongside o1-mini , a smaller , cheaper version . And yes , if you ’ re steeped in AI rumors : this is , in fact , the extremely hyped Strawberry model . For OpenAI , o1 represents a step toward its broader goal of human-like artificial intelligence .
-
-OpenAI is releasing a new model called o1 , the first in a planned series of “ reasoning ” models that have been trained to answer more complex questions , faster than a human can . It ’ s being released alongside o1-mini , a smaller , cheaper version . And yes , if you ’ re steeped in AI rumors : this is , in fact , the extremely hyped Strawberry model . For OpenAI , o1 represents a step toward its broader goal of human-like artificial intelligence . More practically , it does a better job at writing code and solving multistep problems than previous models . But it ’ s also more expensive and slower to use than GPT-4o . OpenAI is calling this release of o1 a “ preview ” to emphasize how nascent it is . ChatGPT Plus and Team users get access to both o1-preview and o1-mini starting today , while Enterprise and Edu users will get access early next week .
-
-More practically , it does a better job at writing code and solving multistep problems than previous models . But it ’ s also more expensive and slower to use than GPT-4o . OpenAI is calling this release of o1 a “ preview ” to emphasize how nascent it is . ChatGPT Plus and Team users get access to both o1-preview and o1-mini starting today , while Enterprise and Edu users will get access early next week . OpenAI says it plans to bring o1-mini access to all the free users of ChatGPT but hasn ’ t set a release date yet . Developer access to o1 is really expensive : In the API , o1-preview is $ 15 per 1 million input tokens , or chunks of text parsed by the model , and $ 60 per 1 million output tokens . For comparison , GPT-4o costs $ 5 per 1 million input tokens and $ 15 per 1 million output tokens .
-
-OpenAI says it plans to bring o1-mini access to all the free users of ChatGPT but hasn ’ t set a release date yet . Developer access to o1 is really expensive : In the API , o1-preview is $ 15 per 1 million input tokens , or chunks of text parsed by the model , and $ 60 per 1 million output tokens . For comparison , GPT-4o costs $ 5 per 1 million input tokens and $ 15 per 1 million output tokens . The training behind o1 is fundamentally different from its predecessors , OpenAI ’ s research lead , Jerry Tworek , tells me , though the company is being vague about the exact details . He says o1 “ has been trained using a completely new optimization algorithm and a new training dataset specifically tailored for it. ” Image : OpenAI OpenAI taught previous GPT models to mimic patterns from its training data .
-
-LLM Answer:
-The text does not mention the CEO of OpenAI. It only discusses the release of a new model called o1 and some details about it, but does not provide information on the company's leadership.
-
-LLM Claim: The text does not mention the CEO of OpenAI.
-Is this claim supported by the context according to bespoke-minicheck? Yes
-
-LLM Claim: It only discusses the release of a new model called o1 and some details about it, but does not provide information on the company's leadership.
-Is this claim supported by the context according to bespoke-minicheck? No
-```
-
-The second claim is unsupported since the text mentions the research lead. 
-
-Another tricky example:
-
-```text
-
-Enter your question or type quit: what sets o1 apart from gpt-4o?
-
-Retrieved chunks: 
-OpenAI says it plans to bring o1-mini access to all the free users of ChatGPT but hasn ’ t set a release date yet . Developer access to o1 is really expensive : In the API , o1-preview is $ 15 per 1 million input tokens , or chunks of text parsed by the model , and $ 60 per 1 million output tokens . For comparison , GPT-4o costs $ 5 per 1 million input tokens and $ 15 per 1 million output tokens . The training behind o1 is fundamentally different from its predecessors , OpenAI ’ s research lead , Jerry Tworek , tells me , though the company is being vague about the exact details . He says o1 “ has been trained using a completely new optimization algorithm and a new training dataset specifically tailored for it. ” Image : OpenAI OpenAI taught previous GPT models to mimic patterns from its training data .
-
-He says OpenAI also tested o1 against a qualifying exam for the International Mathematics Olympiad , and while GPT-4o only correctly solved only 13 percent of problems , o1 scored 83 percent . “ We can ’ t say we solved hallucinations ” In online programming contests known as Codeforces competitions , this new model reached the 89th percentile of participants , and OpenAI claims the next update of this model will perform “ similarly to PhD students on challenging benchmark tasks in physics , chemistry and biology. ” At the same time , o1 is not as capable as GPT-4o in a lot of areas . It doesn ’ t do as well on factual knowledge about the world .
-
-More practically , it does a better job at writing code and solving multistep problems than previous models . But it ’ s also more expensive and slower to use than GPT-4o . OpenAI is calling this release of o1 a “ preview ” to emphasize how nascent it is . ChatGPT Plus and Team users get access to both o1-preview and o1-mini starting today , while Enterprise and Edu users will get access early next week . OpenAI says it plans to bring o1-mini access to all the free users of ChatGPT but hasn ’ t set a release date yet . Developer access to o1 is really expensive : In the API , o1-preview is $ 15 per 1 million input tokens , or chunks of text parsed by the model , and $ 60 per 1 million output tokens . For comparison , GPT-4o costs $ 5 per 1 million input tokens and $ 15 per 1 million output tokens .
-
-“ We can ’ t say we solved hallucinations ” In online programming contests known as Codeforces competitions , this new model reached the 89th percentile of participants , and OpenAI claims the next update of this model will perform “ similarly to PhD students on challenging benchmark tasks in physics , chemistry and biology. ” At the same time , o1 is not as capable as GPT-4o in a lot of areas . It doesn ’ t do as well on factual knowledge about the world . It also doesn ’ t have the ability to browse the web or process files and images . Still , the company believes it represents a brand-new class of capabilities . It was named o1 to indicate “ resetting the counter back to 1. ” “ I ’ m gon na be honest : I think we ’ re terrible at naming , traditionally , ” McGrew says .
-LLM Answer: According to the text, several things set o1 apart from GPT-4o:
-
-* In online programming contests (Codeforces competitions), o1 scored 83% correct solutions compared to GPT-4o's 13%.
-* The training behind o1 is "fundamentally different" from its predecessors, including a completely new optimization algorithm and a new training dataset specifically tailored for it.
-* o1 has been shown to perform similarly to PhD students on challenging benchmark tasks in physics, chemistry, and biology, while GPT-4o does not have this level of performance.
-* However, o1 is also less capable than GPT-4o in some areas, particularly with regard to factual knowledge about the world.
-
-LLM Claim: According to the text, several things set o1 apart from GPT-4o:
-
-* In online programming contests (Codeforces competitions), o1 scored 83% correct solutions compared to GPT-4o's 13%.
-Is this claim supported by the context according to bespoke-minicheck? Yes
-
-LLM Claim: * The training behind o1 is "fundamentally different" from its predecessors, including a completely new optimization algorithm and a new training dataset specifically tailored for it.
-Is this claim supported by the context according to bespoke-minicheck? Yes
-
-LLM Claim: * o1 has been shown to perform similarly to PhD students on challenging benchmark tasks in physics, chemistry, and biology, while GPT-4o does not have this level of performance.
-Is this claim supported by the context according to bespoke-minicheck? No
-
-LLM Claim: * However, o1 is also less capable than GPT-4o in some areas, particularly with regard to factual knowledge about the world.
-Is this claim supported by the context according to bespoke-minicheck? Yes
-```
-
-We see that the third claim "* o1 has been shown to perform similarly to PhD students on challenging benchmark tasks in physics, chemistry, and biology, while GPT-4o does not have this level of performance." is not supported by the context. This is because the context only mentions that o1 "is claimed to perform" which is different from "has been shown to perform".
--- a/examples/python-grounded-factuality-rag-check/main.py
+++ b/examples/python-grounded-factuality-rag-check/main.py
-import ollama
-import warnings
-from mattsollamatools import chunker
-from newspaper import Article
-import numpy as np
-from sklearn.neighbors import NearestNeighbors
-import nltk
-
-warnings.filterwarnings(
-    "ignore", category=FutureWarning, module="transformers.tokenization_utils_base"
-)
-nltk.download("punkt_tab", quiet=True)
-
-
-def getArticleText(url):
-    """Gets the text of an article from a URL.
-
-    Often there are a bunch of ads and menus on pages for a news article.
-    This uses newspaper3k to get just the text of just the article.
-    """
-    article = Article(url)
-    article.download()
-    article.parse()
-    return article.text
-
-
-def knn_search(question_embedding, embeddings, k=5):
-    """Performs K-nearest neighbors (KNN) search"""
-    X = np.array(
-        [item["embedding"] for article in embeddings for item in article["embeddings"]]
-    )
-    source_texts = [
-        item["source"] for article in embeddings for item in article["embeddings"]
-    ]
-
-    # Fit a KNN model on the embeddings
-    knn = NearestNeighbors(n_neighbors=k, metric="cosine")
-    knn.fit(X)
-
-    # Find the indices and distances of the k-nearest neighbors.
-    _, indices = knn.kneighbors(question_embedding, n_neighbors=k)
-
-    # Get the indices and source texts of the best matches
-    best_matches = [(indices[0][i], source_texts[indices[0][i]]) for i in range(k)]
-
-    return best_matches
-
-
-def check(document, claim):
-    """Checks if the claim is supported by the document by calling bespoke-minicheck.
-
-    Returns Yes/yes if the claim is supported by the document, No/no otherwise.
-    Support for logits will be added in the future.
-
-    bespoke-minicheck's system prompt is defined as:
-      'Determine whether the provided claim is consistent with the corresponding
-      document. Consistency in this context implies that all information presented in the claim
-      is substantiated by the document. If not, it should be considered inconsistent. Please
-      assess the claim's consistency with the document by responding with either "Yes" or "No".'
-
-    bespoke-minicheck's user prompt is defined as:
-      "Document: {document}\nClaim: {claim}"
-    """
-    prompt = f"Document: {document}\nClaim: {claim}"
-    response = ollama.generate(
-        model="bespoke-minicheck", prompt=prompt, options={"num_predict": 2, "temperature": 0.0}
-    )
-    return response["response"].strip()
-
-
-if __name__ == "__main__":
-    allEmbeddings = []
-    default_url = "https://www.theverge.com/2024/9/12/24242439/openai-o1-model-reasoning-strawberry-chatgpt"
-    user_input = input(
-        "Enter the URL of an article you want to chat with, or press Enter for default example: "
-    )
-    article_url = user_input.strip() if user_input.strip() else default_url
-    article = {}
-    article["embeddings"] = []
-    article["url"] = article_url
-    text = getArticleText(article_url)
-    chunks = chunker(text)
-
-    # Embed (batch) chunks using ollama
-    embeddings = ollama.embed(model="all-minilm", input=chunks)["embeddings"]
-
-    for chunk, embedding in zip(chunks, embeddings):
-        item = {}
-        item["source"] = chunk
-        item["embedding"] = embedding
-        item["sourcelength"] = len(chunk)
-        article["embeddings"].append(item)
-
-    allEmbeddings.append(article)
-
-    print(f"\nLoaded, chunked, and embedded text from {article_url}.\n")
-
-    while True:
-        # Input a question from the user
-        # For example, "Who is the chief research officer?"
-        question = input("Enter your question or type quit: ")
-
-        if question.lower() == "quit":
-            break
-
-        # Embed the user's question using ollama.embed
-        question_embedding = ollama.embed(model="all-minilm", input=question)[
-            "embeddings"
-        ]
-
-        # Perform KNN search to find the best matches (indices and source text)
-        best_matches = knn_search(question_embedding, allEmbeddings, k=4)
-
-        sourcetext = "\n\n".join([source_text for (_, source_text) in best_matches])
-
-        print(f"\nRetrieved chunks: \n{sourcetext}\n")
-
-        # Give the retrieved chunks and question to the chat model
-        system_prompt = f"Only use the following information to answer the question. Do not use anything else: {sourcetext}"
-
-        ollama_response = ollama.generate(
-            model="llama3.2",
-            prompt=question,
-            system=system_prompt,
-            options={"stream": False},
-        )
-
-        answer = ollama_response["response"]
-        print(f"LLM Answer:\n{answer}\n")
-
-        # Check each sentence in the response for grounded factuality
-        if answer:
-            for claim in nltk.sent_tokenize(answer):
-                print(f"LLM Claim: {claim}")
-                print(
-                    f"Is this claim supported by the context according to bespoke-minicheck? {check(sourcetext, claim)}\n"
-                )
--- a/examples/python-grounded-factuality-rag-check/requirements.txt
+++ b/examples/python-grounded-factuality-rag-check/requirements.txt
-ollama
-lxml==5.3.0
-lxml_html_clean==0.2.2
-mattsollamatools==0.0.25
-newspaper3k==0.2.8
-nltk==3.9.1
-numpy==1.26.4
-scikit-learn==1.5.2
\ No newline at end of file
--- a/examples/python-grounded-factuality-simple-check/main.py
+++ b/examples/python-grounded-factuality-simple-check/main.py
-"""Simple example to demonstrate how to use the bespoke-minicheck model."""
-
-import ollama
-
-# NOTE: ollama must be running for this to work, start the ollama app or run `ollama serve`
-
-
-def check(document, claim):
-    """Checks if the claim is supported by the document by calling bespoke-minicheck.
-
-    Returns Yes/yes if the claim is supported by the document, No/no otherwise.
-    Support for logits will be added in the future.
-
-    bespoke-minicheck's system prompt is defined as:
-      'Determine whether the provided claim is consistent with the corresponding
-      document. Consistency in this context implies that all information presented in the claim
-      is substantiated by the document. If not, it should be considered inconsistent. Please
-      assess the claim's consistency with the document by responding with either "Yes" or "No".'
-
-    bespoke-minicheck's user prompt is defined as:
-      "Document: {document}\nClaim: {claim}"
-    """
-    prompt = f"Document: {document}\nClaim: {claim}"
-    response = ollama.generate(
-        model="bespoke-minicheck", prompt=prompt, options={"num_predict": 2, "temperature": 0.0}
-    )
-    return response["response"].strip()
-
-
-def get_user_input(prompt):
-    user_input = input(prompt)
-    if not user_input:
-        exit()
-    print()
-    return user_input
-
-
-def main():
-    while True:
-        # Get a document from the user (e.g. "Ryan likes running and biking.")
-        document = get_user_input("Enter a document: ")
-        # Get a claim from the user (e.g. "Ryan likes to run.")
-        claim = get_user_input("Enter a claim: ")
-        # Check if the claim is supported by the document
-        grounded_factuality_check = check(document, claim)
-        print(
-            f"Is the claim supported by the document according to bespoke-minicheck? {grounded_factuality_check}"
-        )
-        print("\n\n")
-
-
-if __name__ == "__main__":
-    main()
--- a/examples/python-grounded-factuality-simple-check/readme.md
+++ b/examples/python-grounded-factuality-simple-check/readme.md
-# Simple Bespoke-Minicheck Example
-
-`bespoke-minicheck` is a model for checking if a claim is supported by a document. It is used through the **generate** endpoint, which is called in this example with a `prompt` that includes the expected formatting of the user input. 
-
-## Running the Example
-
-1. Ensure you have the `bespoke-minicheck` model installed:
-
-   ```bash
-   ollama pull bespoke-minicheck
-   ```
-
-2. Install the dependencies:
-
-   ```bash
-   pip install -r requirements.txt
-   ```
-
-3. Run the program:
-
-   ```bash
-   python main.py
-   ```
-
-4. Enter a document and a claim when prompted:
-
-   ```bash
-   Enter a document: Roses are red.
-
-   Enter a claim: Roses are blue. 
-   ```
-
-   The claim and document are then given to the `bespoke-minicheck` as inputs, which then generates a response (Yes or No) on whether the claim is supported by the document.
-
-   ```bash
-   Is the claim supported by the document according to bespoke-minicheck? No
-   ```
-
-## More Examples
-
-Document ([source](https://en.wikipedia.org/wiki/Apple_I)): 
-> The Apple Computer 1 (Apple-1[a]), later known predominantly as the Apple I(written with a Roman numeral),[b] is an 8-bit motherboard-only personal computer designed by Steve Wozniak[5][6] and released by the Apple Computer Company (now Apple Inc.) in 1976. The company was initially formed to sell the Apple I – its first product – and would later become the world's largest technology company.[7] The idea of starting a company and selling the computer came from Wozniak's friend and Apple co-founder Steve Jobs.[8][9] One of the main innovations of the Apple I was that it included video display terminal circuitry on its circuit board, allowing it to connect to a low-cost composite video monitor or television, instead of an expensive computer terminal, compared to most existing computers at the time.
-
-Claim: 
->The Apple I is a 16-bit computer.
-
-Expected output:
->Is the claim supported by the document according to bespoke-minicheck? **No**
-
-Claim: 
->Apple was originally called the Apple Computer Company.
-
-Expected output:
->Is the claim supported by the document according to bespoke-minicheck? **Yes**
--- a/examples/python-grounded-factuality-simple-check/requirements.txt
+++ b/examples/python-grounded-factuality-simple-check/requirements.txt
-ollama
--- a/examples/python-json-datagenerator/predefinedschema.py
+++ b/examples/python-json-datagenerator/predefinedschema.py
-import requests
-import json
-import random
-
-model = "llama3.2"
-template = {
-  "firstName": "",
-  "lastName": "",
-  "address": {
-    "street": "",
-    "city": "",
-    "state": "",
-    "zipCode": ""
-  },
-  "phoneNumber": ""
-}
-
-prompt = f"generate one realistically believable sample data set of a persons first name, last name, address in the US, and  phone number. \nUse the following template: {json.dumps(template)}."
-
-data = {
-    "prompt": prompt,
-    "model": model,
-    "format": "json",
-    "stream": False,
-    "options": {"temperature": 2.5, "top_p": 0.99, "top_k": 100},
-}
-
-print(f"Generating a sample user")
-response = requests.post("http://localhost:11434/api/generate", json=data, stream=False)
-json_data = json.loads(response.text)
-print(json.dumps(json.loads(json_data["response"]), indent=2))
--- a/examples/python-json-datagenerator/randomaddresses.py
+++ b/examples/python-json-datagenerator/randomaddresses.py
-import requests
-import json
-import random
-
-countries = [
-    "United States",
-    "United Kingdom",
-    "the Netherlands",
-    "Germany",
-    "Mexico",
-    "Canada",
-    "France",
-]
-country = random.choice(countries)
-model = "llama3.2"
-
-prompt = f"generate one realistically believable sample data set of a persons first name, last name, address in {country}, and phone number. Do not use common names. Respond using JSON. Key names should have no backslashes, values should use plain ascii with no special characters."
-
-data = {
-    "prompt": prompt,
-    "model": model,
-    "format": "json",
-    "stream": False,
-    "options": {"temperature": 2.5, "top_p": 0.99, "top_k": 100},
-}
-
-print(f"Generating a sample user in {country}")
-response = requests.post("http://localhost:11434/api/generate", json=data, stream=False)
-json_data = json.loads(response.text)
-
-print(json.dumps(json.loads(json_data["response"]), indent=2))
--- a/examples/python-json-datagenerator/readme.md
+++ b/examples/python-json-datagenerator/readme.md
-# JSON Output Example
-
-![llmjson 2023-11-10 15_31_31](https://github.com/ollama/ollama/assets/633681/e599d986-9b4a-4118-81a4-4cfe7e22da25)
-
-There are two python scripts in this example. `randomaddresses.py` generates random addresses from different countries. `predefinedschema.py` sets a template for the model to fill in.
-
-## Running the Example
-
-1. Ensure you have the `llama3.2` model installed:
-
-   ```bash
-   ollama pull llama3.2
-   ```
-
-2. Install the Python Requirements.
-
-   ```bash
-   pip install -r requirements.txt
-   ```
-
-3. Run the Random Addresses example:
-
-   ```bash
-   python randomaddresses.py
-   ```
-
-4. Run the Predefined Schema example:
-
-   ```bash
-   python predefinedschema.py
-   ```
-
-## Review the Code
-
-Both programs are basically the same, with a different prompt for each, demonstrating two different ideas. The key part of getting JSON out of a model is to state in the prompt or system prompt that it should respond using JSON, and specifying the `format` as `json` in the data body.
-
-```python
-prompt = f"generate one realistically believable sample data set of a persons first name, last name, address in {country}, and  phone number. Do not use common names. Respond using JSON. Key names should with no backslashes, values should use plain ascii with no special characters."
-
-data = {
-    "prompt": prompt,
-    "model": model,
-    "format": "json",
-    "stream": False,
-    "options": {"temperature": 2.5, "top_p": 0.99, "top_k": 100},
-}
-```
-
-When running `randomaddresses.py` you will see that the schema changes and adapts to the chosen country.
-
-In `predefinedschema.py`, a template has been specified in the prompt as well. It's been defined as JSON and then dumped into the prompt string to make it easier to work with.
-
-Both examples turn streaming off so that we end up with the completed JSON all at once. We need to convert the `response.text` to JSON so that when we output it as a string we can set the indent spacing to make the output easy to read.
-
-```python
-response = requests.post("http://localhost:11434/api/generate", json=data, stream=False)
-json_data = json.loads(response.text)
-
-print(json.dumps(json.loads(json_data["response"]), indent=2))
-```
--- a/examples/python-json-datagenerator/requirements.txt
+++ b/examples/python-json-datagenerator/requirements.txt
-Requests==2.31.0
--- a/examples/python-loganalysis/Modelfile
+++ b/examples/python-loganalysis/Modelfile
-FROM codebooga:latest
-
-SYSTEM """
-You are a log file analyzer. You will receive a set of lines from a log file for some software application, find the errors and other interesting aspects of the logs, and explain them so a new user can understand what they mean. If there are any steps they can do to resolve them, list the steps in your answer.
-"""
-
-PARAMETER temperature 0.3
-
--- a/examples/python-loganalysis/loganalysis.py
+++ b/examples/python-loganalysis/loganalysis.py
-import sys
-import re
-import requests
-import json
-
-# prelines and postlines represent the number of lines of context to include in the output around the error
-prelines = 10
-postlines = 10
-
-def find_errors_in_log_file():
-  if len(sys.argv) < 2:
-    print("Usage: python loganalysis.py <filename>")
-    return
-
-  log_file_path = sys.argv[1]
-  with open(log_file_path, 'r') as log_file:
-    log_lines = log_file.readlines()
-
-  error_logs = []
-  for i, line in enumerate(log_lines):
-      if "error" in line.lower():
-          start_index = max(0, i - prelines)
-          end_index = min(len(log_lines), i + postlines + 1)
-          error_logs.extend(log_lines[start_index:end_index])
-
-  return error_logs
-
-error_logs = find_errors_in_log_file()
-
-data = {
-  "prompt": "\n".join(error_logs), 
-  "model": "mattw/loganalyzer"
-}
-
-response = requests.post("http://localhost:11434/api/generate", json=data, stream=True)
-for line in response.iter_lines():
-  if line:
-    json_data = json.loads(line)
-    if json_data['done'] == False:
-      print(json_data['response'], end='', flush=True)
-
--- a/examples/python-loganalysis/logtest.logfile
+++ b/examples/python-loganalysis/logtest.logfile
-2023-11-10 07:17:40 /docker-entrypoint.sh: /docker-entrypoint.d/ is not empty, will attempt to perform configuration
-2023-11-10 07:17:40 /docker-entrypoint.sh: Looking for shell scripts in /docker-entrypoint.d/
-2023-11-10 07:17:40 /docker-entrypoint.sh: Launching /docker-entrypoint.d/10-listen-on-ipv6-by-default.sh
-2023-11-10 07:17:40 10-listen-on-ipv6-by-default.sh: info: Getting the checksum of /etc/nginx/conf.d/default.conf
-2023-11-10 07:17:40 10-listen-on-ipv6-by-default.sh: info: Enabled listen on IPv6 in /etc/nginx/conf.d/default.conf
-2023-11-10 07:17:40 /docker-entrypoint.sh: Sourcing /docker-entrypoint.d/15-local-resolvers.envsh
-2023-11-10 07:17:40 /docker-entrypoint.sh: Launching /docker-entrypoint.d/20-envsubst-on-templates.sh
-2023-11-10 07:17:40 /docker-entrypoint.sh: Launching /docker-entrypoint.d/30-tune-worker-processes.sh
-2023-11-10 07:17:40 /docker-entrypoint.sh: Configuration complete; ready for start up
-2023-11-10 07:17:40 2023/11/10 13:17:40 [notice] 1#1: using the "epoll" event method
-2023-11-10 07:17:40 2023/11/10 13:17:40 [notice] 1#1: nginx/1.25.3
-2023-11-10 07:17:40 2023/11/10 13:17:40 [notice] 1#1: built by gcc 12.2.0 (Debian 12.2.0-14) 
-2023-11-10 07:17:40 2023/11/10 13:17:40 [notice] 1#1: OS: Linux 6.4.16-linuxkit
-2023-11-10 07:17:40 2023/11/10 13:17:40 [notice] 1#1: getrlimit(RLIMIT_NOFILE): 1048576:1048576
-2023-11-10 07:17:40 2023/11/10 13:17:40 [notice] 1#1: start worker processes
-2023-11-10 07:17:40 2023/11/10 13:17:40 [notice] 1#1: start worker process 29
-2023-11-10 07:17:40 2023/11/10 13:17:40 [notice] 1#1: start worker process 30
-2023-11-10 07:17:40 2023/11/10 13:17:40 [notice] 1#1: start worker process 31
-2023-11-10 07:17:40 2023/11/10 13:17:40 [notice] 1#1: start worker process 32
-2023-11-10 07:17:40 2023/11/10 13:17:40 [notice] 1#1: start worker process 33
-2023-11-10 07:17:40 2023/11/10 13:17:40 [notice] 1#1: start worker process 34
-2023-11-10 07:17:40 2023/11/10 13:17:40 [notice] 1#1: start worker process 35
-2023-11-10 07:17:40 2023/11/10 13:17:40 [notice] 1#1: start worker process 36
-2023-11-10 07:17:40 2023/11/10 13:17:40 [notice] 1#1: start worker process 37
-2023-11-10 07:17:40 2023/11/10 13:17:40 [notice] 1#1: start worker process 38
-2023-11-10 07:17:44 192.168.65.1 - - [10/Nov/2023:13:17:43 +0000] "GET / HTTP/1.1" 200 615 "-" "Mozilla/5.0 (Macintosh; Intel Mac OS X 10_15_7) AppleWebKit/537.36 (KHTML, like Gecko) Chrome/119.0.0.0 Safari/537.36" "-"
-2023-11-10 07:17:44 2023/11/10 13:17:44 [error] 29#29: *1 open() "/usr/share/nginx/html/favicon.ico" failed (2: No such file or directory), client: 192.168.65.1, server: localhost, request: "GET /favicon.ico HTTP/1.1", host: "localhost:8080", referrer: "http://localhost:8080/"
-2023-11-10 07:17:44 192.168.65.1 - - [10/Nov/2023:13:17:44 +0000] "GET /favicon.ico HTTP/1.1" 404 555 "http://localhost:8080/" "Mozilla/5.0 (Macintosh; Intel Mac OS X 10_15_7) AppleWebKit/537.36 (KHTML, like Gecko) Chrome/119.0.0.0 Safari/537.36" "-"
-2023-11-10 07:17:50 2023/11/10 13:17:50 [error] 29#29: *1 open() "/usr/share/nginx/html/ahstat" failed (2: No such file or directory), client: 192.168.65.1, server: localhost, request: "GET /ahstat HTTP/1.1", host: "localhost:8080"
-2023-11-10 07:17:50 192.168.65.1 - - [10/Nov/2023:13:17:50 +0000] "GET /ahstat HTTP/1.1" 404 555 "-" "Mozilla/5.0 (Macintosh; Intel Mac OS X 10_15_7) AppleWebKit/537.36 (KHTML, like Gecko) Chrome/119.0.0.0 Safari/537.36" "-"
-2023-11-10 07:18:53 2023/11/10 13:18:53 [error] 29#29: *1 open() "/usr/share/nginx/html/ahstat" failed (2: No such file or directory), client: 192.168.65.1, server: localhost, request: "GET /ahstat HTTP/1.1", host: "localhost:8080"
-2023-11-10 07:18:53 192.168.65.1 - - [10/Nov/2023:13:18:53 +0000] "GET /ahstat HTTP/1.1" 404 555 "-" "Mozilla/5.0 (Macintosh; Intel Mac OS X 10_15_7) AppleWebKit/537.36 (KHTML, like Gecko) Chrome/119.0.0.0 Safari/537.36" "-"
--- a/examples/python-loganalysis/readme.md
+++ b/examples/python-loganalysis/readme.md
-# Log Analysis example
-
-![loganalyzer 2023-11-10 08_53_29](https://github.com/ollama/ollama/assets/633681/ad30f1fc-321f-4953-8914-e30e24db9921)
-
-This example shows one possible way to create a log file analyzer. It uses the model **mattw/loganalyzer** which is based on **codebooga**, a 34b parameter model.
-
-To use it, run:
-
-`python loganalysis.py <logfile>`
-
-You can try this with the `logtest.logfile` file included in this directory.
-
-## Running the Example
-
-1. Ensure you have the `mattw/loganalyzer` model installed:
-
-   ```bash
-   ollama pull mattw/loganalyzer
-   ```
-
-2. Install the Python Requirements.
-
-   ```bash
-   python3 -m venv .venv
-   source .venv/bin/activate
-   pip install -r requirements.txt
-   ```
-
-3. Run the example:
-
-   ```bash
-   python loganalysis.py logtest.logfile
-   ```
-
-## Review the code
-
-The first part of this example is a Modelfile that takes `codebooga` and applies a new System Prompt:
-
-```plaintext
-SYSTEM """
-You are a log file analyzer. You will receive a set of lines from a log file for some software application, find the errors and other interesting aspects of the logs, and explain them so a new user can understand what they mean. If there are any steps they can do to resolve them, list the steps in your answer.
-"""
-```
-
-This model is available at https://ollama.com/mattw/loganalyzer. You can customize it and add to your own namespace using the command `ollama create <namespace/modelname> -f <path-to-modelfile>` then `ollama push <namespace/modelname>`.
-
-Then loganalysis.py scans all the lines in the given log file and searches for the word 'error'. When the word is found, the 10 lines before and after are set as the prompt for a call to the Generate API.
-
-```python
-data = {
-  "prompt": "\n".join(error_logs),
-  "model": "mattw/loganalyzer"
-}
-```
-
-Finally, the streamed output is parsed and the response field in the output is printed to the line.
-
-```python
-response = requests.post("http://localhost:11434/api/generate", json=data, stream=True)
-for line in response.iter_lines():
-  if line:
-    json_data = json.loads(line)
-    if json_data['done'] == False:
-      print(json_data['response'], end='')
-
-```
-
-## Next Steps
-
-There is a lot more that can be done here. This is a simple way to detect errors, looking for the word error. Perhaps it would be interesting to find anomalous activity in the logs. It could be interesting to create embeddings for each line and compare them, looking for similar lines. Or look into applying Levenshtein Distance algorithms to find similar lines to help identify the anomalous lines.
-
-Try different models and different prompts to analyze the data. You could consider adding retrieval augmented generation (RAG) to this to help understand newer log formats.
--- a/examples/python-loganalysis/requirements.txt
+++ b/examples/python-loganalysis/requirements.txt
-Requests>=2.32.3