Chaos, complexity, curiosity and database systems. A place where research meets industry
Welcome
Passionately curious about Data, Databases and Systems Complexity. Data is ubiquitous, the database universe is dichotomous (structured and unstructured), expanding and complex. Find my Database Research at SQLToolkit.co.uk . Microsoft Data Platform MVP
"The important thing is not to stop questioning. Curiosity has its own reason for existing" Einstein
"The important thing is not to stop questioning. Curiosity has its own reason for existing" Einstein
Thursday, 9 November 2017
Tuesday, 7 November 2017
Innovative Designs for Innovative Thinking
These Spheres in Seattle will be home of a botanical garden, waterfalls, a river and tree-house like. They are disruptive, pioneering and futuristic to let Amazonians break free for innovative thinking. What an amazingly cool idea.
Thursday, 2 November 2017
PASS Summit 2017 Day 2 Keynote
I attended the Day 2 Keynote at PASS Summit presented by
Rimma Nehme on Globally Distributed Databases Made Simple. This was an amazing
presentation. It was presented seamlessly, explaining the technicalities of
CosmosDB and how the globally distributed database works from the ground up.
Rimma raised the question, do we need another database? Databases need to meet the data needs for today and the future. Data is global, with large volumes of data being created every
60 seconds, which are continually growing and data is interconnected. The balance is shifting in the type of data and we need to have data
globally next to users for processing, meaning the architecture needs to be different.
CosmosDB was originally call Project Florence and was named as such because it is the place where
the renaissance began. It was built in the cloud database for global
distribution, with a fully resource governed stack and schema agnostic service. A single system image is used for all globally distributed resources.
The resource model may have a database account / database that may span clusters and regions. The database is scaled out in terms of containers. It is designed to scale throughput and storage independently. There are two parts to the design. The physical system design is:
The partitioning system design is:
The design is to enable elastically scalable storage, throughput, anywhere, anytime.
Resource governance cannot be an afterthought. The request unit/sec (RU) is the normalized currency.
There are 5 well-defined consistency models in Azure Cosmos
DB with clear trade offs: strong;
bounded-stateless; sessions; consistent prefix and eventual.
There is native support for multiple data models with more coming in the future.
The talk continued to cover how indexing works in depth and the key points to remember about Cosmosdb are:
The talk concluded with a great quote “It is not the
strongest of the species that survives, nor the most intelligent that survives.
It is the one that is most adaptable to change.”
The slides can be downloaded.
Wednesday, 1 November 2017
PASS Summit 2017 Day 1 Keynote
I attended PASS Summit 2017 which was my second year of
attendance. I enjoyed the conference enormously. It is enjoyable being immersed in data and being with people who are enthusiastic in the field.
The Day 1 Keynote "Microsoft for the Modern Data
Estate" was presented by Rohan Kumar. Data is driving transformation. Data, Cloud and AI are the three most disruptive trends of our time. The modern data estate, enables simplicity and common sense.
It takes any data from any source, structured or unstructured data and large or small data. The modern data estate provides a seamless infrastructure between on premises, private and public
cloud, enabling a hybrid set up that hides the dichotomy of these disparate
systems. Seamless flexibly and a choice of engines.
New features in SQL Server 2017
There are many changes to SQL Server 2017. SQL Server 2017 has industry leading performance and security
now on Linux and Docker. The key engine changes
- Support for graph data and queries
- Advanced Machine Learning with R and Python
- Native T-SQL scoring
- Adaptive Query Processing and Automation Plan Correction
SQL
Server 2017 will enable deployment in seconds on Linux and Windows containers
and has special pricing for SQL Server on Linux and Red Hat Enterprise Linux.
New Features Azure SQL Database
Azure SQL Database offers intelligent DBaaS, privacy and trust, seamless and compatibility and competitive TCO. There is seamless migration to the cloud with the cloud first approach
breading faster innovations. The list of changes presented
Azure Data Factory now provides a managed environment for SQL Server Integration Services (SSIS) packages and easily move your SSIS workloads to cloud.
There was the announcement made for a new tool called, Microsoft SQL Operations Studio, a free lightweight modern data operations tools for SQL everywhere.
These are but some of the changes coming to the products.
Thursday, 19 October 2017
Machines that learn to see and move: The future of artificial intelligence
I attended
the Institute for Mathematical Innovation (IMI) public lecture by Professor Andrew Blake, Research
Director at The Alan Turing Institute on 18 October. Professor Blake is a pioneer in the
development of algorithms that make it possible for computers to behave as
seeing machines. Before joining the Institute in 2015, Professor
Blake held the position of Microsoft Distinguished Scientist and Laboratory
Director at the Microsoft Research Lab in Cambridge, and he has been on the
faculty at Oxford University. He is a part of a new startup FiveAI.
The session abstract:
Neural networks have taken the world of
computing in general and artificial intelligence (AI) in particular by storm.
But in the future, AI will need to revisit these generative
models which are used to make predictions. There are several reasons for this –
system robustness, precision issues, transparency, and the high cost of
labelling data.
This is particularly true for perceptual AI, needed for
autonomous vehicles, where the need for simulators and the need to confront
novel situations, will demand further development of generative, probabilistic
models.
He talked about
the empirical detector and generative model. At the moment it is the era of deep
learning and neural networks, that sit within the empirical detector area. A black box area of big data and optimal predictive power. The generative
model is analysis by synthesis and comes with an ‘explanation’, like a model.
It starts with a hypothesis, typically
probabilistic. Professor Blake
believes the generative model will come back as perceptual models need this. This is
This was a very insightful lecture and very interesting
to see the mention of analysis by synthesis.
- to simulate labelled data
- for data fusion - to increase reliability
- to make detailed interpretations
- for online simulation - to explain hard to read situations
Monday, 16 October 2017
Agilience Authority Index
I came across
the Agilience Authority Index placing me in the top 250 for SQL Server.
The
Agilience Authority Index shows
how influential you are and looks at your twitter profile. Agilience state your influence is
more than your audience, your influence is your recognized expertise on a topic. Your
profile on agilience.com shows your main topics of influence based on the
Agilience Authority Index.
PhD Thesis
My PhD thesis is now available online.
Holt, Victoria (2017). A Study into Best Practices and Procedures used in the Management of Database Systems. PhD thesis The Open University.
Subscribe to:
Posts (Atom)










