<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
  <channel>
    <title>CS and Maths Tutorials: Bringing it Together</title>
    <description>Tutorials for students at the interface of Computer Science and Mathematics: Bringing it Together Focuses on themes that unite our domains.
</description>
    <link>http://opendsi.cc/tutorials//</link>
    <atom:link href="http://opendsi.cc/tutorials//feed.xml" rel="self" type="application/rss+xml"/>
    <pubDate>Mon, 30 Nov 2015 10:27:07 +0000</pubDate>
    <lastBuildDate>Mon, 30 Nov 2015 10:27:07 +0000</lastBuildDate>
    <generator>Jekyll v2.4.0</generator>
    
      <item>
        <title>Question and Answer</title>
        <description>&lt;p&gt;This session will be a question and answer session about every thing you’ve been through in this first semester, including data hide meetings, Daniel Bernhardt’s talk and the general scope and direction of the course.&lt;/p&gt;

&lt;p&gt;As material for further discussion please read through the &lt;a href=&quot;http://www.dcs.shef.ac.uk/intranet/teaching/public/tutorials/level1/firstyeartutorials.pdf&quot;&gt;&lt;em&gt;Computer Science First Year’s Tutorial
Booklet&lt;/em&gt;&lt;/a&gt;
on &lt;em&gt;assignments and feedback (Semester 1, week 7)&lt;/em&gt; and &lt;em&gt;review of progress (Semester 1, week 9)&lt;/em&gt;. Think about the different forms of
assessment you receive. How do DCS and SoMaS differ in their ways of
providing you feedback? How does the teaching philosophy in DCS and SoMaS differ?&lt;/p&gt;
</description>
        <pubDate>Tue, 01 Dec 2015 16:00:00 +0000</pubDate>
        <link>http://opendsi.cc/tutorials//2015/12/01/question-and-answer.html</link>
        <guid isPermaLink="true">http://opendsi.cc/tutorials//2015/12/01/question-and-answer.html</guid>
        
        
      </item>
    
      <item>
        <title>Probability Modelling: Language and Physics</title>
        <description>&lt;p&gt;For this week’s meeting Daniel Bernhardt from Facebook will present on spelling at Facebook. The meeting will be followed by a “&lt;a href=&quot;http://opendsi.cc/datahide//2015/11/17/facebook-spelling-data-ethics.html&quot;&gt;Data Hide&lt;/a&gt;” meeting at which Daniel will also present.&lt;/p&gt;
</description>
        <pubDate>Tue, 17 Nov 2015 16:00:00 +0000</pubDate>
        <link>http://opendsi.cc/tutorials//2015/11/17/daniel-bernhardt-visit.html</link>
        <guid isPermaLink="true">http://opendsi.cc/tutorials//2015/11/17/daniel-bernhardt-visit.html</guid>
        
        
      </item>
    
      <item>
        <title>Probability Modelling: Language and Physics</title>
        <description>&lt;p&gt;In two week’s time we will have Daniel Bernhardt from Facebook, London presenting to us at the Data Hide. He will spend the tutorial session beforehand discussing the details of spell checkers at Facebook. &lt;/p&gt;

&lt;p&gt;To prepare for the 17th this we we will have a discussion of &lt;a href=&quot;http://en.wikipedia.org/wiki/N-gram&quot;&gt;&lt;em&gt;n-grams&lt;/em&gt;&lt;/a&gt; and how
to model language, and starting thinking more about the spell checking
problem that Daniel will talk about. We’ll have a simple review of
probability and how to relate marginals and conditional probabilities
and think about how probability can be used to represent sequences. In
particular you should refresh yourself on the sum rule of probability
and the product rule of probability.&lt;/p&gt;

&lt;p&gt;Think about this type of modelling and how it relates to Navier Stokes. In particular what are the abstractions in Navier Stokes? What are the abstactions in &lt;em&gt;n&lt;/em&gt;-gram models and how they are applied to language?&lt;/p&gt;

&lt;p&gt;Which model would you think of as “Higher Level” and which model is a “Lower Level” model.&lt;/p&gt;

&lt;p&gt;Navier Stokes is often thought of as a “Physical model” can &lt;em&gt;n&lt;/em&gt;-gram models be thought of in this way? If so how why? If not why not?&lt;/p&gt;

</description>
        <pubDate>Tue, 03 Nov 2015 16:00:00 +0000</pubDate>
        <link>http://opendsi.cc/tutorials//2015/11/03/n-grams.html</link>
        <guid isPermaLink="true">http://opendsi.cc/tutorials//2015/11/03/n-grams.html</guid>
        
        
      </item>
    
      <item>
        <title>Mathematics and Computer Science: The Interface</title>
        <description>&lt;p&gt;Take a look at &lt;a href=&quot;http://en.wikipedia.org/wiki/Navier%E2%80%93Stokes_equations&quot;&gt;&lt;em&gt;Navier Stokes equations&lt;/em&gt;&lt;/a&gt;
and find out what they are used for. What are the pure mathematics aspects of the equations and what are the applied mathematics aspects? Where does computation come in?&lt;/p&gt;

&lt;p&gt;Think about the relationship between Computer Science and Maths. In absolute terms we many of the challenges of Computer Science are how to &lt;em&gt;implement&lt;/em&gt; maths on a Computer (e.g. the Turing test, the way a processor works). For Mathematics, particularly pure mathematics, there is a lot of interest in &lt;em&gt;separating&lt;/em&gt; the maths from the implementation. Mathematics can stand alone, outside implementation, emerging purely through abstract thought.&lt;/p&gt;

&lt;p&gt;Have a look at the Wikipedia page for &lt;a href=&quot;http://en.wikipedia.org/wiki/Principia_Mathematica&quot;&gt;&lt;em&gt;Principia Mathematica&lt;/em&gt;&lt;/a&gt; and &lt;a href=&quot;http://en.wikipedia.org/wiki/G%C3%B6del%27s_incompleteness_theorems&quot;&gt;&lt;em&gt;Godel’s incompleteness theorem&lt;/em&gt;&lt;/a&gt; as well as the &lt;a href=&quot;http://en.wikipedia.org/wiki/Barber_paradox&quot;&gt;&lt;em&gt;Barber Paradox&lt;/em&gt;&lt;/a&gt;. How does this sort of pure mathematical thinking effect Computer Science?&lt;/p&gt;

&lt;p&gt;Tutorial will be followed by a &lt;a href=&quot;http://opendsi.cc/datahide//2015/10/20/RDM-ai.html&quot;&gt;“Data Hide”&lt;/a&gt;. If you are staying, you need to sign up for it &lt;a href=&quot;https://www.eventbrite.co.uk/e/the-data-hide-tickets-19038564860&quot;&gt;here&lt;/a&gt;.&lt;/p&gt;

&lt;h2 id=&quot;non-academic-task&quot;&gt;Non-Academic Task&lt;/h2&gt;

&lt;p&gt;Before this tutorial you also need to read through the&lt;a href=&quot;http://www.dcs.shef.ac.uk/intranet/teaching/public/tutorials/level1/firstyeartutorials.pdf&quot;&gt; &lt;em&gt;Computer Science First Year’s Tutorial
Booklet&lt;/em&gt;&lt;/a&gt; on &lt;em&gt;unfair means and plagiarism&lt;/em&gt;. You need to fill in the form provided and prepare to submit your answers to MOLE2. (Semester 1, week 4).&lt;/p&gt;
</description>
        <pubDate>Tue, 20 Oct 2015 16:00:00 +0000</pubDate>
        <link>http://opendsi.cc/tutorials//2015/10/20/navier-stokes.html</link>
        <guid isPermaLink="true">http://opendsi.cc/tutorials//2015/10/20/navier-stokes.html</guid>
        
        
      </item>
    
      <item>
        <title>Computer Science and Math Interface: Background</title>
        <description>&lt;p&gt;Many computer science departments were spun out of mathematics departments in the 1980s (including Sheffield’s!), but there is a historical connection between the two departments that has meant that dual degrees are relatively common.&lt;/p&gt;

&lt;p&gt;In Sheffield, the original main link was between areas such as formal methods and pure mathematics. Verification of code for example. This relates to Godel’s incompleteness theorem and Alan Turing’s “Turing Machine”. The focus was on the computability of numbers.&lt;/p&gt;

&lt;p&gt;Recently the nature of the interface has changed a lot. While areas such as theorem proving and verification of code are still very important, a major modern challenge is the overwhelming amount of data that we are generating. This data is generated as a direct consequence of the success of computers. It is also the responsibility of computer scientists to deal with it.&lt;/p&gt;

&lt;p&gt;As background reading for this tutorial have a look at the links below:&lt;/p&gt;

&lt;ul&gt;
  &lt;li&gt;
    &lt;p&gt;Article in The Guardian’s media network on &lt;a href=&quot;http://www.theguardian.com/media-network/2015/mar/05/digital-oligarchy-algorithms-personal-data&quot;&gt;Digital Oligarchies&lt;/a&gt; in March 2015.&lt;/p&gt;
  &lt;/li&gt;
  &lt;li&gt;
    &lt;p&gt;Article in The Guardian’s Media and Tech Network on &lt;a href=&quot;http://www.theguardian.com/media-network/2015/jun/12/artificial-intelligence-ai-human-computer&quot;&gt;preventing AI becoming creepy&lt;/a&gt;.&lt;/p&gt;
  &lt;/li&gt;
  &lt;li&gt;
    &lt;p&gt;Article in The Guardian’s Media and Tech Network on &lt;a href=&quot;http://www.theguardian.com/media-network/2015/aug/25/africa-benefit-data-science-information&quot;&gt;How Africa can benefit from the data science revolution&lt;/a&gt;&lt;/p&gt;
  &lt;/li&gt;
&lt;/ul&gt;

&lt;p&gt;Consider each of the articles and what the implications are for mathematicians and computer scientists. What does it mean for the future of this interface? Which issues fall within the domain of work in data and which fall outside?&lt;/p&gt;

</description>
        <pubDate>Tue, 06 Oct 2015 00:00:00 +0000</pubDate>
        <link>http://opendsi.cc/tutorials//2015/10/06/data-background.html</link>
        <guid isPermaLink="true">http://opendsi.cc/tutorials//2015/10/06/data-background.html</guid>
        
        
      </item>
    
      <item>
        <title>Welcome to Sheffield!</title>
        <description>&lt;p&gt;This first session will be held at “The Hide” in Sheffield, near the Fire and Police Museum. There will be street food and beer and an opportunity to meet others on your course as well as research students who work at the interface of computer science and mathematics. You will also get a chance to meet some of the staff from the two departments.&lt;/p&gt;

&lt;p&gt;You need to register for the event here:&lt;/p&gt;

&lt;p&gt;&lt;a href=&quot;https://www.eventbrite.com/e/the-data-hide-tickets-18817784500&quot;&gt;Eventbrite Registration&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;If you want to know about future events like this please ‘like’ &lt;a href=&quot;https://www.facebook.com/odsi.sheffield&quot;&gt;our Facebook page&lt;/a&gt; or &lt;a href=&quot;https://twitter.com/datascienceshef&quot;&gt;follow us on Twitter&lt;/a&gt;&lt;/p&gt;
</description>
        <pubDate>Tue, 29 Sep 2015 00:00:00 +0000</pubDate>
        <link>http://opendsi.cc/tutorials//2015/09/29/Welcome-to-Sheffield.html</link>
        <guid isPermaLink="true">http://opendsi.cc/tutorials//2015/09/29/Welcome-to-Sheffield.html</guid>
        
        
      </item>
    
      <item>
        <title>Daniel Bernhardt Visit</title>
        <description>&lt;p&gt;Daniel is visiting from Facebook London where he’s been working for just
over a year. Before that he worked for 4 years at Microsoft Search, and
previous to that he did his PhD in applied machine learning for
emotional gesture recognition and his Undergraduate in Computer Science
(both at the University of Cambridge).&lt;/p&gt;

&lt;p&gt;Daniel Bernhardt visit. We will learn about how Facebook deal with the
challenge of spell correction for the vast range of names that exist.
Daniel will give us a talk followed by a Q&amp;amp;A about anything you care to
have answered!&lt;/p&gt;

&lt;p&gt;Daniel spoke about the spell checking problem, two different types of
spelling errors cognitive errors and typo errors, and how these are
passed into the search engine. Search engine queries take several
hundred milliseconds. Your typed query is passed to the search engine
after spell checking.&lt;/p&gt;

&lt;p&gt;“A System for searching the social Graph” (VLDB 2013)&lt;/p&gt;

&lt;h3 id=&quot;first-approach&quot;&gt;First Approach&lt;/h3&gt;

&lt;p&gt;We talked about the &lt;a href=&quot;http://en.wikipedia.org/wiki/Edit_distance&quot;&gt;&lt;em&gt;edit
distance&lt;/em&gt;&lt;/a&gt; as a way of
comparing correct names (in a dictionary) to incorrect names. This is a
dynamic programming approach for computing the distance of the query
from a ‘correct’ query.&lt;/p&gt;

&lt;p&gt;Through a &lt;a href=&quot;http://en.wikipedia.org/wiki/Hash_table&quot;&gt;&lt;em&gt;hash table&lt;/em&gt;&lt;/a&gt; you
can compute look ups quickly.&lt;/p&gt;

&lt;p&gt;In practice the dictionary could be really massive! 1 billion user
nodes, 240 billion photo nodes, 1.2 trillion posts, 1 trillion edges.&lt;/p&gt;

&lt;p&gt;To build a dictionary of unique names: there are 73 million unique first
name and 67 million unique surnames. 560,800,810 unique en-us names! The
dictionary would be very large and nearly 40% of US full names are
unique!&lt;/p&gt;

&lt;p&gt;Context also needs to be included, you are often spelling friends names,
you don’t want the average of all names!&lt;/p&gt;

&lt;p&gt;Edit distances also don’t encode properly mistakes that are more likely
(for example due to a keyboard layout).&lt;/p&gt;

&lt;h3 id=&quot;actual-approach&quot;&gt;Actual Approach&lt;/h3&gt;

&lt;p&gt;Language model takes any given sequence and gives a probability
associated. n-gram models can be used for modelling language.&lt;/p&gt;

&lt;p&gt;We talked about error models and how they can be derived from query
rewrites. This is when a user types a search and then doesn’t seem to
select any of the returned results. They then change the query and that
new query gives the right result. This gives a lot of valuable
information about what the original query should have been.&lt;/p&gt;

&lt;p&gt;We talked about challenges including doing different languages, and
making general solutions. In particular we talked about the challenge of
&lt;em&gt;domain adaptation&lt;/em&gt; and how a spell checker for one particular domain
may not be appropriate for another domain. We talked about the
importance of data and feedback.&lt;/p&gt;

&lt;p&gt;Daniel talked about how widely machine learning is used across the site,
and how, often, it’s not the learning algorithm itself that’s a problem,
but evaluating the quality of the model or extracting the data for
learning.&lt;/p&gt;

&lt;p&gt;Charles asked about when the system is finished, Daniel said broadly
speaking that they are always looking for continuous improvement. Whilst
the target is to focus on impact, they don’t consider a system done, but
if the system is working well, then they may reprioritise on something
else.&lt;/p&gt;

&lt;p&gt;** Also Non Academic Task 4 **&lt;/p&gt;

&lt;h2 id=&quot;non-academic-4&quot;&gt;Non-Academic 4&lt;/h2&gt;

&lt;p&gt;Read through the&lt;a href=&quot;http://www.dcs.shef.ac.uk/intranet/teaching/public/tutorials/level1/firstyeartutorials.pdf&quot;&gt;&lt;em&gt;Computer Science First Year’s Tutorial
Booklet&lt;/em&gt;&lt;/a&gt;
on &lt;em&gt;exam skills&lt;/em&gt;. Read old exam papers and think about what the
question’s aims are. How to DCS and SoMaS differ in their exam question
style? How are they similar?&lt;/p&gt;

</description>
        <pubDate>Mon, 24 Nov 2014 00:00:00 +0000</pubDate>
        <link>http://opendsi.cc/tutorials//2014/11/24/daniel-bernhardt-visit.html</link>
        <guid isPermaLink="true">http://opendsi.cc/tutorials//2014/11/24/daniel-bernhardt-visit.html</guid>
        
        
      </item>
    
      <item>
        <title>Preparation for Daniel Bernhardt Visit</title>
        <description>&lt;p&gt;Preparation for Daniel’s visit! Let’s go through those rules of
probability once again, and have quick refresher on n-gram models. We
should also try and look at dynamic programming algorithms for trying to
find the most probable path through a sequence.&lt;/p&gt;

</description>
        <pubDate>Mon, 17 Nov 2014 00:00:00 +0000</pubDate>
        <link>http://opendsi.cc/tutorials//2014/11/17/daniel-bernhardt-preparation.html</link>
        <guid isPermaLink="true">http://opendsi.cc/tutorials//2014/11/17/daniel-bernhardt-preparation.html</guid>
        
        
      </item>
    
      <item>
        <title>John Baez Topics 2</title>
        <description>&lt;p&gt;Presentation in pairs or individuals of topics from John Baez’s Google+
feed Week 2.&lt;/p&gt;

&lt;p&gt;Group discussion, lead by Matheus and Islam, about Science, Models and
Machine Learning&lt;/p&gt;

&lt;p&gt;&lt;a href=&quot;http://johncarlosbaez.wordpress.com/2014/09/03/science-models-and-machine-learning/&quot;&gt;http://johncarlosbaez.wordpress.com/2014/09/03/science-models-and-machine-learning/&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;Some of the topics mentioned during the discussion were:&lt;/p&gt;

&lt;p&gt;Software available:&lt;/p&gt;

&lt;p&gt;&lt;a href=&quot;http://scikit-learn.org/stable/&quot;&gt;http://scikit-learn.org/stable/&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;&lt;a href=&quot;https://github.com/SheffieldML/GPy&quot;&gt;https://github.com/SheffieldML/GPy&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;&lt;a href=&quot;http://www.automaticstatistician.com&quot;&gt;http://www.automaticstatistician.com&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;Decision making and rational behaviour:
&lt;a href=&quot;http://www.ted.com/talks/dan_ariely_asks_are_we_in_control_of_our_own_decisions#t-890383&quot;&gt;http://www.ted.com/talks/dan_ariely_asks_are_we_in_control_of_our_own_decisions#t-890383&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;The sampling theory (statistical sampling):&lt;/p&gt;

&lt;p&gt;&lt;a href=&quot;http://www.youtube.com/watch?v=7nr9rQpm2A4&quot;&gt;http://www.youtube.com/watch?v=7nr9rQpm2A4&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;Also: &lt;strong&gt;Non Academic Task 3 (see below)&lt;/strong&gt;&lt;/p&gt;

&lt;h2 id=&quot;non-academic-3&quot;&gt;Non-Academic 3&lt;/h2&gt;

&lt;p&gt;Read through the&lt;a href=&quot;http://www.dcs.shef.ac.uk/intranet/teaching/public/tutorials/level1/firstyeartutorials.pdf&quot;&gt; &lt;em&gt;Computer Science First Year’s Tutorial
Booklet&lt;/em&gt;&lt;/a&gt;
on &lt;em&gt;assignments and feedback&lt;/em&gt;. Think about the different forms of
assessment you receive. How do DCS and SoMaS differ in their ways of
providing you feedback? (Semester 1, week 7)&lt;/p&gt;

</description>
        <pubDate>Mon, 10 Nov 2014 00:00:00 +0000</pubDate>
        <link>http://opendsi.cc/tutorials//john%20baez/non%20academic%203/automatic%20statistician/scikit-learn/github/decision%20making/2014/11/10/john-baez-2.html</link>
        <guid isPermaLink="true">http://opendsi.cc/tutorials//john%20baez/non%20academic%203/automatic%20statistician/scikit-learn/github/decision%20making/2014/11/10/john-baez-2.html</guid>
        
        
        <category>john baez</category>
        
        <category>non academic 3</category>
        
        <category>automatic statistician</category>
        
        <category>scikit-learn</category>
        
        <category>github</category>
        
        <category>decision making</category>
        
      </item>
    
      <item>
        <title>John Baez Topics 1</title>
        <description>&lt;p&gt;John Baez topics Week 1&lt;/p&gt;

&lt;p&gt;Thomas presented on the Quantum Thriller:
&lt;a href=&quot;https://www.youtube.com/watch?v=I7oZAo8hJnU&quot;&gt;&lt;em&gt;https://www.youtube.com/watch?v=I7oZAo8hJnU&lt;/em&gt;&lt;/a&gt;
which led us on to Quantum cryptography. Interesting reading is:&lt;/p&gt;

&lt;p&gt;&lt;a href=&quot;http://en.wikipedia.org/wiki/A_Universe_from_Nothing&quot;&gt;&lt;em&gt;http://en.wikipedia.org/wiki/A_Universe_from_Nothing&lt;/em&gt;&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;&lt;a href=&quot;http://simonsingh.net/books/the-code-book/&quot;&gt;&lt;em&gt;http://simonsingh.net/books/the-code-book/&lt;/em&gt;&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;Presentation in pairs or individuals of topics from John Baez’s Google+
feed Week 1.&lt;/p&gt;

</description>
        <pubDate>Mon, 03 Nov 2014 00:00:00 +0000</pubDate>
        <link>http://opendsi.cc/tutorials//2014/11/03/john-baez-1.html</link>
        <guid isPermaLink="true">http://opendsi.cc/tutorials//2014/11/03/john-baez-1.html</guid>
        
        
      </item>
    
  </channel>
</rss>
